a2-highgpu-2g by Google Cloud Platform
(All-cores)
(Single-core)
Specifications
Server Metadata
Vendor ID | gcp |
Server ID | 1000024 |
Name | a2-highgpu-2g |
Description | Accelerator Optimized: 2 NVIDIA Tesla A100 GPUs, 24 vCPUs, 170GB RAM |
Family | a2 |
Hw Virt | |
Average Time To Start | 57 |
Status | active |
Observed At | 2026-08-06T12:09:15.627504 |
Availability
| REGION / ID | SPOT | ONDEMAND |
|---|---|---|
Council Bluffs (US) / us-central1 | 0.8538 USD/h | 1.479 USD/h |
The Dalles (US) / us-west1 | 0.8538 USD/h | 1.479 USD/h |
Moncks Corner (US) / us-east1 | 0.8538 USD/h | 1.479 USD/h |
Tel Aviv (IL) / me-west1 | 0.976 USD/h | 1.6268 USD/h |
Eemshaven (NL) / europe-west4 | 0.809 USD/h | 1.6281 USD/h |
Las Vegas (US) / us-west4 | 0.8075 USD/h | 1.6656 USD/h |
Salt Lake City (US) / us-west3 | 1.0657 USD/h | 1.7764 USD/h |
Jurong West (SG) / asia-southeast1 | 1.0947 USD/h | 1.8244 USD/h |
Seoul (KR) / asia-northeast3 | 1.1082 USD/h | 1.8961 USD/h |
Tokyo (JP) / asia-northeast1 | 1.1376 USD/h | 1.8961 USD/h |
Processor
vCPUs | 24 |
CPU Allocation | Dedicated |
CPU Cores | 12 |
CPU Speed | 2.2 GHz |
CPU Architecture | x86_64 |
CPU Manufacturer | Intel |
CPU Family | Xeon |
CPU L1D Cache | 32 KiB |
CPU L1D Cache Total | 384 KiB |
CPU L1I Cache | 32 KiB |
CPU L1I Cache Total | 384 KiB |
CPU L2 Cache | 1 MiB |
CPU L2 Cache Total | 12 MiB |
CPU L3 Cache | 39 MiB |
CPU L3 Cache Total | 39 MiB |
CPU Flags | fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ss, ht, syscall, nx, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, xtopology, nonstop_tsc, cpuid, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, abm, 3dnowprefetch, ssbd, ibrs, ibpb, stibp, ibrs_enhanced, fsgsbase, tsc_adjust, bmi1, hle, avx2, smep, bmi2, erms, invpcid, rtm, mpx, avx512f, avx512dq, rdseed, adx, smap, clflushopt, clwb, avx512cd, avx512bw, avx512vl, xsaveopt, xsavec, xgetbv1, xsaves, arat, avx512_vnni, md_clear, arch_capabilities |
Ecpus | 15.8 |
Scalability | 131.67 |
System Resources and Accelerators
| MEMORY | |
|---|---|
Memory Amount | 170 GiB |
Memory Amount Actual | 170 GiB |
| GPU | |
|---|---|
GPU Count | 2 |
GPU Memory Min | 40 GiB |
GPU Memory Total | 80 GiB |
GPU Manufacturer | NVIDIA |
GPU Family | Ampere |
GPU Model | A100 |
GPUs |
|
| STORAGE | |
|---|---|
Storage Size | 0 GB |
Storages |
| NETWORK | |
|---|---|
Inbound Traffic | 0 GB/month |
Outbound Traffic | 0 GB/month |
IPv4 | 0 |
CPU and System Topology
Server Description
An accelerator-optimized platform featuring dual GPUs and high-capacity memory, engineered for top-tier large language model inference and parallel computing workloads.
Google Cloud Platform a2-highgpu-2g is an accelerator-optimized server designed for GPU-intensive workloads. It features two NVIDIA Ampere A100 GPUs with 80 GB of total VRAM, paired with an Intel Xeon x86_64 processor offering 24 dedicated vCPUs (12 physical cores) running at 2.2 GHz. The system includes 170.0 GB of RAM but lacks local storage. Benchmark data shows that while single-core CPU performance is in the bottom 25%, multi-core and memory bandwidth metrics are average. Crucially, the server delivers top-tier performance in the top 10% for large language model (LLM) prompt processing and text generation across medium and large models. This profile aligns with machine learning inference, deep learning, and parallel GPU computing, whereas it is less optimal for sequential CPU-bound tasks.
Economics
Performance
Memory Bandwidth
Compression
OpenSSL
Geekbench Single-Core
Geekbench Multi-Core
Passmark CPU Scores
| BENCHMARK | SCORE |
|---|---|
Mark | 19022 |
Compression | 278622 |
Encryption | 8685 |
Extended Instructions | 18027 |
Floating Point Maths | 34100 |
Integer Maths | 63469 |
Physics | 1952 |
Prime Numbers | 102 |
Single Threaded | 1693 |
String Sorting | 32865 |
Passmark Memory Scores
| BENCHMARK | SCORE |
|---|---|
Memory Mark | 2483 |
Database Operations | 7693 |
Memory Latency | 56 |
Memory Read Cached | 20838 |
Memory Read Uncached | 10561 |
Memory Write | 9923 |
Stress-ng Raw Scores
Stress-ng Relative Multicore Performance
LLM Inference Speed for Prompt Processing
LLM Inference Speed for Text Generation
Static Web Server
Redis
Alternatives
Servers of the Same Family
| INSTANCE | vCPUs | MEMORY | GPUs |
|---|---|---|---|
| a2-highgpu-1g | 12 | 85 GiB | 1 |
| a2-ultragpu-1g | 12 | 170 GiB | 1 |
| a2-ultragpu-2g | 24 | 340 GiB | 2 |
| a2-highgpu-4g | 48 | 340 GiB | 4 |
| a2-ultragpu-4g | 48 | 680 GiB | 4 |
| a2-highgpu-8g | 96 | 680 GiB | 8 |
| a2-ultragpu-8g | 96 | 1360 GiB | 8 |
Similar Servers
| INSTANCE | VENDOR | vCPUs | MEMORY | GPUs |
|---|