logo

g5.16xlarge by Amazon Web Services

g5.16xlarge is a Graphics intensive Gen5 16xlarge server offered by Amazon Web Services with 64 vCPUs, 256 GiB of memory and 1.9 TB of storage. The pricing starts at 0.5733 USD per hour.
64 vCPU
256 GiB Memory
1.9 TB Storage
1 GPU
Spare SCore
24581
(All-cores)
767
(Single-core)

Specifications

Server Metadata

Vendor ID
aws
Name
g5.16xlarge
Description
Graphics intensive Gen5 16xlarge
Family
g5
Hw Virt
Status
active
Observed At
2026-07-25T19:53:22.718403

Availability

REGION / IDSPOTONDEMAND
Ohio (US) / us-east-2
1.109533 USD/h 4.096 USD/h
Oregon (US) / us-west-2
1.647667 USD/h 4.096 USD/h
Northern Virgina (US) / us-east-1
2.09186 USD/h 4.096 USD/h
Stockholm (SE) / eu-north-1
1.14505 USD/h 4.3444 USD/h
Quebec (CA) / ca-central-1
1.09075 USD/h 4.5479 USD/h
Dublin (IE) / eu-west-1
1.2592 USD/h 4.5724 USD/h
Tel Aviv (IL) / il-central-1
2.443 USD/h 4.801 USD/h
Mumbai (IN) / ap-south-1
1.3648 USD/h 4.9185 USD/h
United Arab Emirates / me-central-1
2.0568 USD/h 5.0243 USD/h
Seoul (KR) / ap-northeast-2
1.260033 USD/h 5.0365 USD/h
Frankfurt (DE) / eu-central-1
2.0036 USD/h 5.122 USD/h
London (GB) / eu-west-2
2.28895 USD/h 5.1994 USD/h
Sydney (AU) / ap-southeast-2
1.53215 USD/h 5.3256 USD/h
Jakarta (ID) / ap-southeast-3
3.0284 USD/h 5.7328 USD/h
Tokyo (JP) / ap-northeast-1
1.1138 USD/h 5.9404 USD/h
Hong Kong (HK) / ap-east-1
2.03135 USD/h 6.306 USD/h
Sao Paulo (BR) / sa-east-1
5.7176 USD/h 6.9624 USD/h

Processor

vCPUs
64
Hypervisor
nitro
CPU Allocation
Dedicated
CPU Cores
32
CPU Speed
3.3 GHz
CPU Architecture
x86_64
CPU Manufacturer
AMD
CPU Family
EPYC
CPU Model
7R32
CPU L1D Cache
32 KiB
CPU L1D Cache Total
1 MiB
CPU L1I Cache
32 KiB
CPU L1I Cache Total
1 MiB
CPU L2 Cache
512 KiB
CPU L2 Cache Total
16 MiB
CPU L3 Cache
16 MiB
CPU L3 Cache Total
128 MiB
CPU Flags
fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, mmxext, fxsr_opt, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, nonstop_tsc, cpuid, extd_apicid, aperfmperf, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, sse4_1, sse4_2, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, cmp_legacy, cr8_legacy, abm, sse4a, misalignsse, 3dnowprefetch, topoext, ssbd, ibrs, ibpb, stibp, vmmcall, fsgsbase, bmi1, avx2, smep, bmi2, rdseed, adx, smap, clflushopt, clwb, sha_ni, xsaveopt, xsavec, xgetbv1, clzero, xsaveerptr, rdpru, wbnoinvd, arat, npt, nrip_save, rdpid
Ecpus
32.1
Scalability
100.31

System Resources and Accelerators

MEMORY
Memory Amount
256 GiB
Memory Amount Actual
256 GiB
Memory Generation
DDR4
Memory Speed
3200 Mhz
GPU
GPU Count
1
GPU Memory Min
22 GiB
GPU Memory Total
22 GiB
GPU Manufacturer
NVIDIA
GPU Family
Ampere
GPU Model
A10G
GPUs
  • NVIDIA Ampere NVIDIA A10G (Memory amount: 23028, Firmware version: 590.48.01, BIOS version: 94.02.75.00.01, Clock rate: 1710)
STORAGE
Storage Size
1900 GB
Storage Type
nvme ssd
Storages
  • 1900 GB nvme ssd
NETWORK
Network Speed Baseline
25 Gbps
Network Speed Max
25 Gbps
Network Storage Speed Baseline
16 Gbps
Network Storage Speed Max
16 Gbps
Inbound Traffic
0 GB/month
Outbound Traffic
0 GB/month
IPv4
0

CPU and System Topology

Server Description

An accelerated virtual server combining dedicated AMD EPYC processors, local NVMe storage, and NVIDIA Ampere graphics for demanding machine learning workloads.

GPU AcceleratedGeneral Purpose

Amazon Web Services g5.16xlarge is a graphics-intensive instance featuring 64 dedicated vCPUs from an AMD EPYC 7R32 processor running at 3.3 GHz, 256.0 GB of DDR4 memory, and a baseline network bandwidth of 25 Gbps. It includes a bundled NVIDIA Ampere A10G GPU with 22 GB of VRAM and 1900 GB of local NVMe SSD storage. Performance benchmarks show a contrast between poor single-core CPU metrics and top-tier LLM inference speeds for small and medium models. While PassMark CPU and memory scores are strong, text generation for large 70B models is weak. This configuration offers qualitative cost efficiency by combining dedicated compute, high-speed local storage, and GPU acceleration on the Nitro hypervisor. It is best suited for machine learning inference, ray tracing, and workloads requiring high-density GPU and storage resources.

Economics

Average Price per Region

Prices per Zone

Lowest Prices

Performance

Workload Profiles

2.09Score
Precomputed compound score for Cache Intensive workloads. A weighted average (geometric mean) of benchmark scores compared to their medians: score = ∏ (x_i / m_i)^(w_i / Σw). The score of 1.0 represents a synthetic baseline server with the median performance of each component benchmark; 0.5 means roughly half the performance; and 2.0 means twice the performance of that reference profile. Component weights: 50% Redis RPS (pipeline=1, SET), 20% Redis RPS (pipeline=16, SET), 10% PassMark Memory Mark (composite), 10% Memory bandwidth (read, 16 MB ~ L3), 10% PassMark single-thread CPU. Rationale for component selection: In-memory key-value store workload, mixing direct Redis performance metrics with memory speed and latency benchmarks, and single-core CPU performance profiles.
Component Score Weight Impact
Raw Reference Normalized Target Actual
Redis RPS (pipeline=1, SET) 5,141,109 ops/sec 1,896,415 ops/sec 2.71 50.00% 50.00% +64.60%
Redis RPS (pipeline=16, SET) 34,564,529 ops/sec 11,539,523 ops/sec 3 20.00% 20.00% +24.60%
PassMark Memory Mark (composite) 2,917 2,411 1.21 10.00% 10.00% +1.92%
Memory bandwidth (read, 16 MB ~ L3) 119,992 MB/sec 107,368 MB/sec 1.12 10.00% 10.00% +1.14%
PassMark single-thread CPU 2,135 Mops/s 2,363 Mops/s 0.904 10.00% 10.00% -1.00%

Memory Bandwidth

Compression

OpenSSL

Geekbench Single-Core

Score: 1,332

Geekbench Multi-Core

Score: 14,069

Passmark CPU Scores

BENCHMARKSCORE
Mark
55615
Compression
953024
Encryption
68193
Extended Instructions
56952
Floating Point Maths
128866
Integer Maths
216068
Physics
5371
Prime Numbers
296
Single Threaded
2135
String Sorting
126195

Passmark Memory Scores

BENCHMARKSCORE
Memory Mark
2917
Database Operations
21347
Memory Latency
51
Memory Read Cached
24632
Memory Read Uncached
12768
Memory Write
13485

Stress-ng Raw Scores

Stress-ng Relative Multicore Performance

LLM Inference Speed for Prompt Processing

LLM Inference Speed for Text Generation

Static Web Server

Redis

Alternatives

Servers of the Same Family

INSTANCEvCPUsMEMORYGPUs
g5.xlarge 416 GiB1
g5.2xlarge 832 GiB1
g5.4xlarge 1664 GiB1
g5.8xlarge 32128 GiB1
g5.12xlarge 48192 GiB4
g5.24xlarge 96384 GiB4
g5.48xlarge 192768 GiB8

Similar Servers

INSTANCEVENDORvCPUsMEMORYGPUs

g5.16xlarge FAQs