logo

a2-highgpu-4g by Google Cloud Platform

a2-highgpu-4g is a Accelerator Optimized: 4 NVIDIA Tesla A100 GPUs, 48 vCPUs, 340GB RAM server offered by Google Cloud Platform with 48 vCPUs, 340 GiB of memory and 0 GB of storage. The pricing starts at 1.4709 USD per hour.
48 vCPU
340 GiB Memory
4 GPU
Spare SCore
37929
(All-cores)
1201
(Single-core)

Specifications

Server Metadata

Vendor ID
gcp
Server ID
1000048
Name
a2-highgpu-4g
Description
Accelerator Optimized: 4 NVIDIA Tesla A100 GPUs, 48 vCPUs, 340GB RAM
Family
a2
Hw Virt
Average Time To Start
43
Status
active
Observed At
2026-07-29T01:14:38.068806

Availability

REGION / IDSPOTONDEMAND
Council Bluffs (US) / us-central1
1.5525 USD/h 2.9579 USD/h
The Dalles (US) / us-west1
1.5525 USD/h 2.9579 USD/h
Moncks Corner (US) / us-east1
1.5525 USD/h 2.9579 USD/h
Tel Aviv (IL) / me-west1
1.9519 USD/h 3.2537 USD/h
Eemshaven (NL) / europe-west4
1.4709 USD/h 3.2563 USD/h
Las Vegas (US) / us-west4
1.5381 USD/h 3.3312 USD/h
Salt Lake City (US) / us-west3
2.1315 USD/h 3.5528 USD/h
Jurong West (SG) / asia-southeast1
2.1894 USD/h 3.6488 USD/h
Seoul (KR) / asia-northeast3
2.0147 USD/h 3.7921 USD/h
Tokyo (JP) / asia-northeast1
2.2751 USD/h 3.7921 USD/h

Processor

vCPUs
48
CPU Allocation
Dedicated
CPU Cores
24
CPU Speed
2.2 GHz
CPU Architecture
x86_64
CPU Manufacturer
Intel
CPU Family
Xeon
CPU L1D Cache
32 KiB
CPU L1D Cache Total
768 KiB
CPU L1I Cache
32 KiB
CPU L1I Cache Total
768 KiB
CPU L2 Cache
1 MiB
CPU L2 Cache Total
24 MiB
CPU L3 Cache
39 MiB
CPU L3 Cache Total
39 MiB
CPU Flags
fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ss, ht, syscall, nx, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, xtopology, nonstop_tsc, cpuid, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, abm, 3dnowprefetch, ssbd, ibrs, ibpb, stibp, ibrs_enhanced, fsgsbase, tsc_adjust, bmi1, hle, avx2, smep, bmi2, erms, invpcid, rtm, mpx, avx512f, avx512dq, rdseed, adx, smap, clflushopt, clwb, avx512cd, avx512bw, avx512vl, xsaveopt, xsavec, xgetbv1, xsaves, arat, avx512_vnni, md_clear, arch_capabilities
Ecpus
31.6
Scalability
131.67

System Resources and Accelerators

MEMORY
Memory Amount
340 GiB
Memory Amount Actual
340 GiB
GPU
GPU Count
4
GPU Memory Min
40 GiB
GPU Memory Total
160 GiB
GPU Manufacturer
NVIDIA
GPU Family
Ampere
GPU Model
A100
GPUs
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.45.00.03, Clock rate: 1410)
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.45.00.03, Clock rate: 1410)
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.45.00.03, Clock rate: 1410)
  • NVIDIA Ampere NVIDIA A100-SXM4-40GB (Memory amount: 40960, Firmware version: 590.48.01, BIOS version: 92.00.45.00.03, Clock rate: 1410)
STORAGE
Storage Size
0 GB
Storages
    NETWORK
    Inbound Traffic
    0 GB/month
    Outbound Traffic
    0 GB/month
    IPv4
    0

    CPU and System Topology

    Server Description

    An accelerator-optimized platform featuring four dedicated GPUs and high-capacity memory for demanding parallel computing and large language model inference.

    GPU AcceleratedMemory Optimized

    Google Cloud Platform a2-highgpu-4g is an accelerator-optimized server featuring four NVIDIA Ampere A100 GPUs with 160 GB of total VRAM, paired with 48 vCPUs on an Intel Xeon x86_64 architecture and 340.0 GB of system memory. While single-core CPU benchmarks place this instance in the bottom 25%, its multi-core capabilities and memory bandwidth sit in the middle 50%. The server demonstrates top-tier performance in the top 10% for large language model inference, specifically in prompt processing and text generation for medium and large models. Lacking local storage, this configuration is designed for GPU-intensive workloads, including deep learning, parallel data science tasks, and LLM deployment, where the high-density accelerator configuration justifies the hardware profile.

    Economics

    Average Price per Region

    Prices per Zone

    Lowest Prices

    Performance

    Workload Profiles

    1.41Score
    Precomputed compound score for Cache Intensive workloads. A weighted average (geometric mean) of benchmark scores compared to their medians: score = ∏ (x_i / m_i)^(w_i / Σw). The score of 1.0 represents a synthetic baseline server with the median performance of each component benchmark; 0.5 means roughly half the performance; and 2.0 means twice the performance of that reference profile. Component weights: 50% Redis RPS (pipeline=1, SET), 20% Redis RPS (pipeline=16, SET), 10% PassMark Memory Mark (composite), 10% Memory bandwidth (read, 16 MB ~ L3), 10% PassMark single-thread CPU. Rationale for component selection: In-memory key-value store workload, mixing direct Redis performance metrics with memory speed and latency benchmarks, and single-core CPU performance profiles.
    Component Score Weight Impact
    Raw Reference Normalized Target Actual
    Redis RPS (pipeline=1, SET) 3,221,348 ops/sec 1,896,415 ops/sec 1.7 50.00% 50.00% +30.40%
    Redis RPS (pipeline=16, SET) 19,056,588 ops/sec 11,539,523 ops/sec 1.65 20.00% 20.00% +10.50%
    PassMark Memory Mark (composite) 2,511 2,411 1.04 10.00% 10.00% +0.39%
    Memory bandwidth (read, 16 MB ~ L3) 115,554 MB/sec 107,368 MB/sec 1.08 10.00% 10.00% +0.77%
    PassMark single-thread CPU 1,702 Mops/s 2,363 Mops/s 0.72 10.00% 10.00% -3.23%

    Memory Bandwidth

    Compression

    OpenSSL

    Geekbench Single-Core

    Score: 1,062

    Geekbench Multi-Core

    Score: 10,469

    Passmark CPU Scores

    BENCHMARKSCORE
    Mark
    31203
    Compression
    533487
    Encryption
    15898
    Extended Instructions
    33690
    Floating Point Maths
    68173
    Integer Maths
    126886
    Physics
    2568
    Prime Numbers
    168
    Single Threaded
    1702
    String Sorting
    59439

    Passmark Memory Scores

    BENCHMARKSCORE
    Memory Mark
    2511
    Database Operations
    14784
    Memory Latency
    53
    Memory Read Cached
    20833
    Memory Read Uncached
    9791
    Memory Write
    9432

    Stress-ng Raw Scores

    Stress-ng Relative Multicore Performance

    LLM Inference Speed for Prompt Processing

    LLM Inference Speed for Text Generation

    Static Web Server

    Redis

    Alternatives

    Servers of the Same Family

    INSTANCEvCPUsMEMORYGPUs
    a2-highgpu-1g 1285 GiB1
    a2-ultragpu-1g 12170 GiB1
    a2-highgpu-2g 24170 GiB2
    a2-ultragpu-2g 24340 GiB2
    a2-ultragpu-4g 48680 GiB4
    a2-highgpu-8g 96680 GiB8
    a2-ultragpu-8g 961360 GiB8

    Similar Servers

    INSTANCEVENDORvCPUsMEMORYGPUs

    a2-highgpu-4g FAQs