a2-highgpu-1g by Google Cloud Platform
(All-cores)
(Single-core)
Specifications
Server Metadata
Vendor ID | gcp |
Server ID | 1000012 |
Name | a2-highgpu-1g |
Description | Accelerator Optimized: 1 NVIDIA Tesla A100 GPU, 12 vCPUs, 85GB RAM |
Family | a2 |
Hw Virt | |
Average Time To Start | 26 |
Status | active |
Observed At | 2026-07-12T09:46:12.943631 |
Availability
| REGION / ID | SPOT | ONDEMAND |
|---|---|---|
Council Bluffs (US) / us-central1 | 0.3881 USD/h | 0.7395 USD/h |
The Dalles (US) / us-west1 | 0.3881 USD/h | 0.7395 USD/h |
Moncks Corner (US) / us-east1 | 0.3881 USD/h | 0.7395 USD/h |
Tel Aviv (IL) / me-west1 | 0.4695 USD/h | 0.8134 USD/h |
Eemshaven (NL) / europe-west4 | 0.3677 USD/h | 0.8141 USD/h |
Las Vegas (US) / us-west4 | 0.3845 USD/h | 0.8328 USD/h |
Salt Lake City (US) / us-west3 | 0.5329 USD/h | 0.8882 USD/h |
Jurong West (SG) / asia-southeast1 | 0.5252 USD/h | 0.9122 USD/h |
Seoul (KR) / asia-northeast3 | 0.4798 USD/h | 0.948 USD/h |
Tokyo (JP) / asia-northeast1 | 0.5476 USD/h | 0.948 USD/h |
Processor
vCPUs | 12 |
CPU Allocation | Dedicated |
CPU Cores | 6 |
CPU Speed | 2.2 GHz |
CPU Architecture | x86_64 |
CPU Manufacturer | Intel |
CPU Family | Xeon |
CPU L1D Cache | 32 KiB |
CPU L1D Cache Total | 192 KiB |
CPU L1I Cache | 32 KiB |
CPU L1I Cache Total | 192 KiB |
CPU L2 Cache | 1 MiB |
CPU L2 Cache Total | 6 MiB |
CPU L3 Cache | 39 MiB |
CPU L3 Cache Total | 39 MiB |
CPU Flags | fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ss, ht, syscall, nx, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, xtopology, nonstop_tsc, cpuid, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, abm, 3dnowprefetch, ssbd, ibrs, ibpb, stibp, ibrs_enhanced, fsgsbase, tsc_adjust, bmi1, hle, avx2, smep, bmi2, erms, invpcid, rtm, mpx, avx512f, avx512dq, rdseed, adx, smap, clflushopt, clwb, avx512cd, avx512bw, avx512vl, xsaveopt, xsavec, xgetbv1, xsaves, arat, avx512_vnni, md_clear, arch_capabilities |
Ecpus | 7.9 |
Scalability | 131.67 |
System Resources and Accelerators
| MEMORY | |
|---|---|
Memory Amount | 85 GiB |
Memory Amount Actual | 85 GiB |
| GPU | |
|---|---|
GPU Count | 1 |
GPU Memory Min | 40 GiB |
GPU Memory Total | 40 GiB |
GPU Manufacturer | NVIDIA |
GPU Family | Ampere |
GPU Model | A100 |
GPUs |
|
| STORAGE | |
|---|---|
Storage Size | 0 GB |
Storages |
| NETWORK | |
|---|---|
Inbound Traffic | 0 GB/month |
Outbound Traffic | 0 GB/month |
IPv4 | 0 |
CPU and System Topology
Server Description
An accelerator-optimized instance pairing an NVIDIA A100 GPU with dedicated Intel Xeon processors for high-throughput machine learning inference.
Google Cloud Platform a2-highgpu-1g is an accelerator-optimized server designed for GPU-intensive workloads. It features a dedicated Intel Xeon CPU with 12 vCPUs, 6 physical cores, and a 2.2 GHz clock speed, paired with 85 GB of system memory. The primary hardware highlight is a bundled NVIDIA Ampere A100 GPU with 40 GB of VRAM. While single-core CPU benchmarks place this instance in the bottom 25% and multi-core performance sits in the middle 50%, its GPU capabilities deliver top-tier performance for large language model inference, particularly with 135M and 7B parameter models. Local storage is not included. This instance is highly cost-effective for machine learning, deep learning, and data science tasks that offload processing to the GPU, though it is less suited for CPU-bound or storage-heavy workloads.
Economics
Performance
Workload Profiles
Memory Bandwidth
Compression
OpenSSL
Geekbench Single-Core
Geekbench Multi-Core
Passmark CPU Scores
| BENCHMARK | SCORE |
|---|---|
Mark | 10529 |
Compression | 141009 |
Encryption | 4354 |
Extended Instructions | 8926 |
Floating Point Maths | 17048 |
Integer Maths | 31697 |
Physics | 1217 |
Prime Numbers | 64 |
Single Threaded | 1698 |
String Sorting | 16898 |
Passmark Memory Scores
| BENCHMARK | SCORE |
|---|---|
Memory Mark | 2333 |
Database Operations | 3886 |
Memory Latency | 55 |
Memory Read Cached | 20650 |
Memory Read Uncached | 9599 |
Memory Write | 9200 |
Stress-ng Raw Scores
Stress-ng Relative Multicore Performance
LLM Inference Speed for Prompt Processing
LLM Inference Speed for Text Generation
Static Web Server
Redis
Alternatives
Servers of the Same Family
| INSTANCE | vCPUs | MEMORY | GPUs |
|---|---|---|---|
| a2-ultragpu-1g | 12 | 170 GiB | 1 |
| a2-highgpu-2g | 24 | 170 GiB | 2 |
| a2-ultragpu-2g | 24 | 340 GiB | 2 |
| a2-highgpu-4g | 48 | 340 GiB | 4 |
| a2-ultragpu-4g | 48 | 680 GiB | 4 |
| a2-highgpu-8g | 96 | 680 GiB | 8 |
| a2-ultragpu-8g | 96 | 1360 GiB | 8 |
Similar Servers
| INSTANCE | VENDOR | vCPUs | MEMORY | GPUs |
|---|