a2-highgpu-8g by Google Cloud Platform
(All-cores)
(Single-core)
Specifications
Server Metadata
Vendor ID | gcp |
Server ID | 1000096 |
Name | a2-highgpu-8g |
Description | Accelerator Optimized: 8 NVIDIA Tesla A100 GPUs, 96 vCPUs, 680GB RAM |
Family | a2 |
Hw Virt | |
Average Time To Start | 78 |
Status | active |
Observed At | 2026-08-06T12:09:15.627638 |
Availability
| REGION / ID | SPOT | ONDEMAND |
|---|---|---|
Council Bluffs (US) / us-central1 | 3.4153 USD/h | 5.9158 USD/h |
The Dalles (US) / us-west1 | 3.4153 USD/h | 5.9158 USD/h |
Moncks Corner (US) / us-east1 | 3.4153 USD/h | 5.9158 USD/h |
Tel Aviv (IL) / me-west1 | 3.9038 USD/h | 6.5074 USD/h |
Eemshaven (NL) / europe-west4 | 3.2361 USD/h | 6.5125 USD/h |
Las Vegas (US) / us-west4 | 3.2298 USD/h | 6.6624 USD/h |
Salt Lake City (US) / us-west3 | 4.2629 USD/h | 7.1056 USD/h |
Jurong West (SG) / asia-southeast1 | 4.3789 USD/h | 7.2976 USD/h |
Seoul (KR) / asia-northeast3 | 4.4326 USD/h | 7.5842 USD/h |
Tokyo (JP) / asia-northeast1 | 4.5502 USD/h | 7.5842 USD/h |
Processor
vCPUs | 96 |
CPU Allocation | Dedicated |
CPU Cores | 48 |
CPU Speed | 2.2 GHz |
CPU Architecture | x86_64 |
CPU Manufacturer | Intel |
CPU Family | Xeon |
CPU L1D Cache | 32 KiB |
CPU L1D Cache Total | 2 MiB |
CPU L1I Cache | 32 KiB |
CPU L1I Cache Total | 2 MiB |
CPU L2 Cache | 1 MiB |
CPU L2 Cache Total | 48 MiB |
CPU L3 Cache | 39 MiB |
CPU L3 Cache Total | 77 MiB |
CPU Flags | fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ss, ht, syscall, nx, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, xtopology, nonstop_tsc, cpuid, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, abm, 3dnowprefetch, ssbd, ibrs, ibpb, stibp, ibrs_enhanced, fsgsbase, tsc_adjust, bmi1, hle, avx2, smep, bmi2, erms, invpcid, rtm, mpx, avx512f, avx512dq, rdseed, adx, smap, clflushopt, clwb, avx512cd, avx512bw, avx512vl, xsaveopt, xsavec, xgetbv1, xsaves, arat, avx512_vnni, md_clear, arch_capabilities |
Ecpus | 63.1 |
Scalability | 131.46 |
System Resources and Accelerators
| MEMORY | |
|---|---|
Memory Amount | 680 GiB |
Memory Amount Actual | 680 GiB |
| GPU | |
|---|---|
GPU Count | 8 |
GPU Memory Min | 40 GiB |
GPU Memory Total | 320 GiB |
GPU Manufacturer | NVIDIA |
GPU Family | Ampere |
GPU Model | A100 |
GPUs |
|
| STORAGE | |
|---|---|
Storage Size | 0 GB |
Storages |
| NETWORK | |
|---|---|
Inbound Traffic | 0 GB/month |
Outbound Traffic | 0 GB/month |
IPv4 | 0 |
CPU and System Topology
Server Description
An accelerator-optimized platform featuring eight dedicated GPUs and massive memory bandwidth for high-throughput parallel computing and large language model inference.
Google Cloud Platform a2-highgpu-8g is an accelerator-optimized server designed for highly parallelized workloads. It features eight NVIDIA Ampere A100 GPUs with 320 GB of total VRAM, paired with 96 dedicated Intel Xeon vCPUs (48 physical cores) running at 2.2 GHz and 680.0 GB of system memory. The server does not include local storage. Performance benchmarks demonstrate top-tier capabilities in large language model inference and strong multi-core CPU execution, though single-core performance remains weak. This hardware profile makes the instance highly suitable for deep learning, LLM prompt processing, and ray tracing, while representing a specialized resource-density tradeoff that is less efficient for single-threaded applications.
Economics
Performance
Memory Bandwidth
Compression
OpenSSL
Geekbench Single-Core
Geekbench Multi-Core
Passmark CPU Scores
| BENCHMARK | SCORE |
|---|---|
Mark | 49728 |
Compression | 1059158 |
Encryption | 31839 |
Extended Instructions | 67287 |
Floating Point Maths | 136204 |
Integer Maths | 253339 |
Physics | 5048 |
Prime Numbers | 333 |
Single Threaded | 1699 |
String Sorting | 118795 |
Passmark Memory Scores
| BENCHMARK | SCORE |
|---|---|
Memory Mark | 2509 |
Database Operations | 20510 |
Memory Latency | 54 |
Memory Read Cached | 20819 |
Memory Read Uncached | 9785 |
Memory Write | 9390 |
Stress-ng Raw Scores
Stress-ng Relative Multicore Performance
LLM Inference Speed for Prompt Processing
LLM Inference Speed for Text Generation
Static Web Server
Redis
Alternatives
Servers of the Same Family
| INSTANCE | vCPUs | MEMORY | GPUs |
|---|---|---|---|
| a2-highgpu-1g | 12 | 85 GiB | 1 |
| a2-ultragpu-1g | 12 | 170 GiB | 1 |
| a2-highgpu-2g | 24 | 170 GiB | 2 |
| a2-ultragpu-2g | 24 | 340 GiB | 2 |
| a2-highgpu-4g | 48 | 340 GiB | 4 |
| a2-ultragpu-4g | 48 | 680 GiB | 4 |
| a2-ultragpu-8g | 96 | 1360 GiB | 8 |
Similar Servers
| INSTANCE | VENDOR | vCPUs | MEMORY | GPUs |
|---|