logo

a2-highgpu-8g by Google Cloud Platform

a2-highgpu-8g is a Accelerator Optimized: 8 NVIDIA Tesla A100 GPUs, 96 vCPUs, 680GB RAM server offered by Google Cloud Platform with 96 vCPUs, 680 GiB of memory and 0 GB of storage. The pricing starts at 3.2298 USD per hour.
96 vCPU
680 GiB Memory
8 GPU
Spare SCore
75800
(All-cores)
1201
(Single-core)

Specifications

Server Metadata

Vendor ID
gcp
Server ID
1000096
Name
a2-highgpu-8g
Description
Accelerator Optimized: 8 NVIDIA Tesla A100 GPUs, 96 vCPUs, 680GB RAM
Family
a2
Hw Virt
Average Time To Start
78
Status
active
Observed At
2026-08-06T12:09:15.627638

Availability

REGION / IDSPOTONDEMAND
Council Bluffs (US) / us-central1
3.4153 USD/h 5.9158 USD/h
The Dalles (US) / us-west1
3.4153 USD/h 5.9158 USD/h
Moncks Corner (US) / us-east1
3.4153 USD/h 5.9158 USD/h
Tel Aviv (IL) / me-west1
3.9038 USD/h 6.5074 USD/h
Eemshaven (NL) / europe-west4
3.2361 USD/h 6.5125 USD/h
Las Vegas (US) / us-west4
3.2298 USD/h 6.6624 USD/h
Salt Lake City (US) / us-west3
4.2629 USD/h 7.1056 USD/h
Jurong West (SG) / asia-southeast1
4.3789 USD/h 7.2976 USD/h
Seoul (KR) / asia-northeast3
4.4326 USD/h 7.5842 USD/h
Tokyo (JP) / asia-northeast1
4.5502 USD/h 7.5842 USD/h

Processor

vCPUs
96
CPU Allocation
Dedicated
CPU Cores
48
CPU Speed
2.2 GHz
CPU Architecture
x86_64
CPU Manufacturer
Intel
CPU Family
Xeon
CPU L1D Cache
32 KiB
CPU L1D Cache Total
2 MiB
CPU L1I Cache
32 KiB
CPU L1I Cache Total
2 MiB
CPU L2 Cache
1 MiB
CPU L2 Cache Total
48 MiB
CPU L3 Cache
39 MiB
CPU L3 Cache Total
77 MiB
CPU Flags
fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ss, ht, syscall, nx, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, xtopology, nonstop_tsc, cpuid, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, abm, 3dnowprefetch, ssbd, ibrs, ibpb, stibp, ibrs_enhanced, fsgsbase, tsc_adjust, bmi1, hle, avx2, smep, bmi2, erms, invpcid, rtm, mpx, avx512f, avx512dq, rdseed, adx, smap, clflushopt, clwb, avx512cd, avx512bw, avx512vl, xsaveopt, xsavec, xgetbv1, xsaves, arat, avx512_vnni, md_clear, arch_capabilities
Ecpus
63.1
Scalability
131.46

System Resources and Accelerators

MEMORY
Memory Amount
680 GiB
Memory Amount Actual
680 GiB
GPU
GPU Count
8
GPU Memory Min
40 GiB
GPU Memory Total
320 GiB
GPU Manufacturer
NVIDIA
GPU Family
Ampere
GPU Model
A100
GPUs
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.81.00.01, Clock rate: 1410)
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.81.00.01, Clock rate: 1410)
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.81.00.01, Clock rate: 1410)
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.81.00.01, Clock rate: 1410)
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.81.00.01, Clock rate: 1410)
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.81.00.01, Clock rate: 1410)
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.81.00.01, Clock rate: 1410)
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.81.00.01, Clock rate: 1410)
STORAGE
Storage Size
0 GB
Storages
    NETWORK
    Inbound Traffic
    0 GB/month
    Outbound Traffic
    0 GB/month
    IPv4
    0

    CPU and System Topology

    Server Description

    An accelerator-optimized platform featuring eight dedicated GPUs and massive memory bandwidth for high-throughput parallel computing and large language model inference.

    GPU AcceleratedMemory OptimizedCompute Optimized

    Google Cloud Platform a2-highgpu-8g is an accelerator-optimized server designed for highly parallelized workloads. It features eight NVIDIA Ampere A100 GPUs with 320 GB of total VRAM, paired with 96 dedicated Intel Xeon vCPUs (48 physical cores) running at 2.2 GHz and 680.0 GB of system memory. The server does not include local storage. Performance benchmarks demonstrate top-tier capabilities in large language model inference and strong multi-core CPU execution, though single-core performance remains weak. This hardware profile makes the instance highly suitable for deep learning, LLM prompt processing, and ray tracing, while representing a specialized resource-density tradeoff that is less efficient for single-threaded applications.

    Economics

    Average Price per Region

    Prices per Zone

    Lowest Prices

    Performance

    Workload Profiles

    2.53Score
    Precomputed compound score for Cache Intensive workloads. A weighted average (geometric mean) of benchmark scores compared to their medians: score = ∏ (x_i / m_i)^(w_i / Σw). The score of 1.0 represents a synthetic baseline server with the median performance of each component benchmark; 0.5 means roughly half the performance; and 2.0 means twice the performance of that reference profile. Component weights: 50% Redis RPS (pipeline=1, SET), 20% Redis RPS (pipeline=16, SET), 10% PassMark Memory Mark (composite), 10% Memory bandwidth (read, 16 MB ~ L3), 10% PassMark single-thread CPU. Rationale for component selection: In-memory key-value store workload, mixing direct Redis performance metrics with memory speed and latency benchmarks, and single-core CPU performance profiles.
    Component Score Weight Impact
    Raw Reference Normalized Target Actual
    Redis RPS (pipeline=1, SET) 6,771,021 ops/sec 1,910,660 ops/sec 3.54 50.00% 50.00% +88.10%
    Redis RPS (pipeline=16, SET) 39,929,924 ops/sec 11,608,287 ops/sec 3.44 20.00% 20.00% +28.00%
    PassMark Memory Mark (composite) 2,509 2,416 1.04 10.00% 10.00% +0.39%
    Memory bandwidth (read, 16 MB ~ L3) 231,594 MB/sec 107,896 MB/sec 2.15 10.00% 10.00% +7.96%
    PassMark single-thread CPU 1,699 Mops/s 2,364 Mops/s 0.718 10.00% 10.00% -3.26%

    Memory Bandwidth

    Compression

    OpenSSL

    Geekbench Single-Core

    Score: 1,063

    Geekbench Multi-Core

    Score: 11,752

    Passmark CPU Scores

    BENCHMARKSCORE
    Mark
    49728
    Compression
    1059158
    Encryption
    31839
    Extended Instructions
    67287
    Floating Point Maths
    136204
    Integer Maths
    253339
    Physics
    5048
    Prime Numbers
    333
    Single Threaded
    1699
    String Sorting
    118795

    Passmark Memory Scores

    BENCHMARKSCORE
    Memory Mark
    2509
    Database Operations
    20510
    Memory Latency
    54
    Memory Read Cached
    20819
    Memory Read Uncached
    9785
    Memory Write
    9390

    Stress-ng Raw Scores

    Stress-ng Relative Multicore Performance

    LLM Inference Speed for Prompt Processing

    LLM Inference Speed for Text Generation

    Static Web Server

    Redis

    Alternatives

    Servers of the Same Family

    INSTANCEvCPUsMEMORYGPUs
    a2-highgpu-1g 1285 GiB1
    a2-ultragpu-1g 12170 GiB1
    a2-highgpu-2g 24170 GiB2
    a2-ultragpu-2g 24340 GiB2
    a2-highgpu-4g 48340 GiB4
    a2-ultragpu-4g 48680 GiB4
    a2-ultragpu-8g 961360 GiB8

    Similar Servers

    INSTANCEVENDORvCPUsMEMORYGPUs

    a2-highgpu-8g FAQs