a3-highgpu-2g by Google Cloud Platform
(All-cores)
(Single-core)
Specifications
Server Metadata
Vendor ID | gcp |
Server ID | 1720362 |
Name | a3-highgpu-2g |
Description | Accelerator Optimized: 2 NVIDIA H100 GPU, 52 vCPUs, 468GB RAM |
Family | a3 |
Hw Virt | |
Average Time To Start | 35 |
Status | active |
Observed At | 2026-10-04T03:14:28.352558Z |
Availability
| REGION / ID | SPOT | ONDEMAND |
|---|---|---|
Council Bluffs (US) / us-central1 | 13.2406 USD/h | 22.1225 USD/h |
The Dalles (US) / us-west1 | 13.2406 USD/h | 22.1225 USD/h |
Ashburn (US) / us-east4 | 13.2833 USD/h | 22.1389 USD/h |
Mumbai (IN) / asia-south1 | 13.7203 USD/h | 23.0337 USD/h |
St. Ghislain (BE) / europe-west1 | 13.8643 USD/h | 24.3425 USD/h |
Hamina (FI) / europe-north1 | 2.2846 USD/h | 24.3589 USD/h |
Columbus (US) / us-east5 | 13.562 USD/h | 24.8958 USD/h |
Las Vegas (US) / us-west4 | 12.3757 USD/h | 24.9122 USD/h |
London (GB) / europe-west2 | 3.0972 USD/h | 25.2295 USD/h |
Changhua County (TW) / asia-east1 | 15.3539 USD/h | 25.5897 USD/h |
Paris (FR) / europe-west9 | 8.9092 USD/h | 25.6621 USD/h |
Frankfurt (DE) / europe-west3 | 15.6638 USD/h | 26.1078 USD/h |
Delhi (IN) / asia-south2 | 7.764 USD/h | 26.5712 USD/h |
Sydney (AU) / australia-southeast1 | 6.639 USD/h | 27.6696 USD/h |
Eemshaven (NL) / europe-west4 | 16.8741 USD/h | 28.1351 USD/h |
Seoul (KR) / asia-northeast3 | 11.4294 USD/h | 28.4123 USD/h |
Jurong West (SG) / asia-southeast1 | 17.1399 USD/h | 28.5696 USD/h |
Melbourne (AU) / australia-southeast2 | 2.7045 USD/h | 28.7675 USD/h |
Tokyo (JP) / asia-northeast1 | 6.384 USD/h | 31.6608 USD/h |
Processor
vCPUs | 52 |
CPU Allocation | Dedicated |
CPU Cores | 26 |
CPU Speed | 2.7 GHz |
CPU Architecture | x86_64 |
CPU Manufacturer | Intel |
CPU Family | Xeon |
CPU Model | 8481C |
CPU L1D Cache | 48 KiB |
CPU L1D Cache Total | 1 MiB |
CPU L1I Cache | 32 KiB |
CPU L1I Cache Total | 832 KiB |
CPU L2 Cache | 2 MiB |
CPU L2 Cache Total | 52 MiB |
CPU L3 Cache | 105 MiB |
CPU L3 Cache Total | 105 MiB |
CPU Flags | fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ss, ht, syscall, nx, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, xtopology, nonstop_tsc, cpuid, tsc_known_freq, pni, pclmulqdq, vmx, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, abm, 3dnowprefetch, ssbd, ibrs, ibpb, stibp, ibrs_enhanced, tpr_shadow, flexpriority, ept, vpid, ept_ad, fsgsbase, tsc_adjust, bmi1, avx2, smep, bmi2, erms, invpcid, rtm, avx512f, avx512dq, rdseed, adx, smap, avx512ifma, clflushopt, clwb, avx512cd, sha_ni, avx512bw, avx512vl, xsaveopt, xsavec, xgetbv1, xsaves, avx_vnni, avx512_bf16, arat, vnmi, avx512vbmi, umip, avx512_vbmi2, gfni, vaes, vpclmulqdq, avx512_vnni, avx512_bitalg, avx512_vpopcntdq, la57, rdpid, cldemote, movdiri, movdir64b, fsrm, md_clear, serialize, tsxldtrk, amx_bf16, avx512_fp16, amx_tile, amx_int8, arch_capabilities |
Ecpus | 22.7 |
Scalability | 43.65 |
System Resources and Accelerators
| MEMORY | |
|---|---|
Memory Amount | 468 GiB |
Memory Amount Actual | 468 GiB |
| GPU | |
|---|---|
GPU Count | 2 |
GPU Memory Min | 80 GiB |
GPU Memory Total | 160 GiB |
GPU Manufacturer | NVIDIA |
GPU Family | Hopper |
GPU Model | H100 |
GPUs |
|
| STORAGE | |
|---|---|
Storage Size | 1608 GB |
Storage Type | nvme ssd |
Storages |
|
| NETWORK | |
|---|---|
Inbound Traffic | 0 GB/month |
Outbound Traffic | 0 GB/month |
IPv4 | 0 |
CPU and System Topology
Server Description
An accelerator-optimized server featuring dual NVIDIA H100 GPUs and high-speed NVMe storage for demanding machine learning and inference workloads.
Google Cloud Platform a3-highgpu-2g is an accelerator-optimized server designed for high-compute and machine learning workloads. It features two NVIDIA H100 Hopper GPUs with 160 GB of total VRAM, paired with 52 vCPUs from an Intel Xeon 8481C processor running at 2.7 GHz. The system includes 468 GB of RAM and a 1608 GB local NVMe SSD. Benchmarks demonstrate top-tier performance in LLM inference, with prompt processing speeds of 5644.37 tokens per second on a 7B model, alongside strong memory bandwidth and multi-threaded decompression capabilities. This server is optimized for deep learning, large language model inference, and high-performance database operations.
Economics
Performance
Memory Bandwidth
Compression
OpenSSL
Geekbench Single-Core
Geekbench Multi-Core
Passmark CPU Scores
| BENCHMARK | SCORE |
|---|---|
Mark | 54314 |
Compression | 716947 |
Encryption | 43985 |
Extended Instructions | 46057 |
Floating Point Maths | 142147 |
Integer Maths | 199432 |
Physics | 6503 |
Prime Numbers | 349 |
Single Threaded | 2720 |
String Sorting | 97370 |
Passmark Memory Scores
| BENCHMARK | SCORE |
|---|---|
Memory Mark | 2230 |
Database Operations | 23460 |
Memory Latency | 74 |
Memory Read Cached | 25230 |
Memory Read Uncached | 9444 |
Memory Write | 8322 |
Stress-ng Raw Scores
Stress-ng Relative Multicore Performance
LLM Inference Speed for Prompt Processing
LLM Inference Speed for Text Generation
Static Web Server
Redis
Alternatives
Servers of the Same Family
| INSTANCE | vCPUs | MEMORY | GPUs |
|---|---|---|---|
| a3-highgpu-1g | 26 | 234 GiB | 1 |
| a3-highgpu-4g | 104 | 936 GiB | 4 |
| a3-highgpu-8g | 208 | 1872 GiB | 8 |
| a3-megagpu-8g | 208 | 1872 GiB | 8 |
| a3-edgegpu-8g | 208 | 1872 GiB | 8 |
| a3-edgegpu-8g-nolssd | 208 | 1872 GiB | 8 |
| a3-ultragpu-8g | 224 | 2952 GiB | 8 |
Similar Servers
| INSTANCE | VENDOR | vCPUs | MEMORY | GPUs |
|---|