g5.16xlarge by Amazon Web Services

g5.16xlarge is a Graphics intensive Gen5 16xlarge server offered by Amazon Web Services with 64 vCPUs, 256 GiB of memory and 1.9 TB of storage. The pricing starts at 0.4801 USD per hour.
64 vCPU
256 GiB Memory
1.9 TB Storage
1 GPU
Spare SCore
24,581
(All-cores)
767
(Single-core)

Specifications

Server Metadata

Vendor ID
aws
Name
g5.16xlarge
Description
Graphics intensive Gen5 16xlarge
Family
g5
Hw Virt
Average Time To Start
11
Status
active
Observed At
2026-09-08T20:54:13.156789

Availability

REGION / IDSPOTONDEMAND
Oregon (US) / us-west-2
1.303033 USD/h 4.096 USD/h
Ohio (US) / us-east-2
1.9862 USD/h 4.096 USD/h
Northern Virgina (US) / us-east-1
2.4364 USD/h 4.096 USD/h
Aragón (ES) / eu-south-2
4.12365 USD/h 4.3159 USD/h
Stockholm (SE) / eu-north-1
0.72085 USD/h 4.3444 USD/h
Quebec (CA) / ca-central-1
1.6867 USD/h 4.5479 USD/h
Dublin (IE) / eu-west-1
1.5766 USD/h 4.5724 USD/h
Tel Aviv (IL) / il-central-1
1.945033 USD/h 4.801 USD/h
Mumbai (IN) / ap-south-1
1.73485 USD/h 4.9185 USD/h
United Arab Emirates / me-central-1
2.0568 USD/h 5.0243 USD/h
Seoul (KR) / ap-northeast-2
2.343467 USD/h 5.0365 USD/h
Frankfurt (DE) / eu-central-1
1.9861 USD/h 5.122 USD/h
Paris (FR) / eu-west-3
2.69165 USD/h 5.1994 USD/h
London (GB) / eu-west-2
3.39245 USD/h 5.1994 USD/h
Sydney (AU) / ap-southeast-2
1.32 USD/h 5.3256 USD/h
Jakarta (ID) / ap-southeast-3
2.8133 USD/h 5.7328 USD/h
Tokyo (JP) / ap-northeast-1
2.7046 USD/h 5.9404 USD/h
Hong Kong (HK) / ap-east-1
1.4415 USD/h 6.306 USD/h
Sao Paulo (BR) / sa-east-1
5.9081 USD/h 6.9624 USD/h

Processor

vCPUs
64
Hypervisor
nitro
CPU Allocation
Dedicated
CPU Cores
32
CPU Speed
3.3 GHz
CPU Architecture
x86_64
CPU Manufacturer
AMD
CPU Family
EPYC
CPU Model
7R32
CPU L1D Cache
32 KiB
CPU L1D Cache Total
1 MiB
CPU L1I Cache
32 KiB
CPU L1I Cache Total
1 MiB
CPU L2 Cache
512 KiB
CPU L2 Cache Total
16 MiB
CPU L3 Cache
16 MiB
CPU L3 Cache Total
128 MiB
CPU Flags
fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, mmxext, fxsr_opt, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, nonstop_tsc, cpuid, extd_apicid, aperfmperf, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, sse4_1, sse4_2, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, cmp_legacy, cr8_legacy, abm, sse4a, misalignsse, 3dnowprefetch, topoext, ssbd, ibrs, ibpb, stibp, vmmcall, fsgsbase, bmi1, avx2, smep, bmi2, rdseed, adx, smap, clflushopt, clwb, sha_ni, xsaveopt, xsavec, xgetbv1, clzero, xsaveerptr, rdpru, wbnoinvd, arat, npt, nrip_save, rdpid
Ecpus
32.1
Scalability
100.31

System Resources and Accelerators

MEMORY
Memory Amount
256 GiB
Memory Amount Actual
256 GiB
Memory Generation
DDR4
Memory Speed
3200 Mhz
GPU
GPU Count
1
GPU Memory Min
22 GiB
GPU Memory Total
22 GiB
GPU Manufacturer
NVIDIA
GPU Family
Ampere
GPU Model
A10G
GPUs
  • NVIDIA Ampere NVIDIA A10G (Memory amount: 23028, Firmware version: 590.48.01, BIOS version: 94.02.75.00.01, Clock rate: 1710)
STORAGE
Storage Size
1900 GB
Storage Type
nvme ssd
Storages
  • 1900 GB nvme ssd
NETWORK
Network Speed Baseline
25 Gbps
Network Speed Max
25 Gbps
Network Storage Speed Baseline
16 Gbps
Network Storage Speed Max
16 Gbps
Inbound Traffic
0 GB/month
Outbound Traffic
0 GB/month
IPv4
0

CPU and System Topology

Server Description

An accelerated virtual server combining dedicated AMD EPYC processors, local NVMe storage, and NVIDIA Ampere graphics for demanding machine learning workloads.

GPU AcceleratedGeneral Purpose

Amazon Web Services g5.16xlarge is a graphics-intensive instance featuring 64 dedicated vCPUs from an AMD EPYC 7R32 processor running at 3.3 GHz, 256.0 GB of DDR4 memory, and a baseline network bandwidth of 25 Gbps. It includes a bundled NVIDIA Ampere A10G GPU with 22 GB of VRAM and 1900 GB of local NVMe SSD storage. Performance benchmarks show a contrast between poor single-core CPU metrics and top-tier LLM inference speeds for small and medium models. While PassMark CPU and memory scores are strong, text generation for large 70B models is weak. This configuration offers qualitative cost efficiency by combining dedicated compute, high-speed local storage, and GPU acceleration on the Nitro hypervisor. It is best suited for machine learning inference, ray tracing, and workloads requiring high-density GPU and storage resources.

Economics

Average Price per Region

Prices per Zone

Lowest Prices

Performance

Workload Profiles

2.07Score
Precomputed compound score for Cache Intensive workloads. A weighted average (geometric mean) of benchmark scores compared to their medians: score = ∏ (x_i / m_i)^(w_i / Σw). The score of 1.0 represents a synthetic baseline server with the median performance of each component benchmark; 0.5 means roughly half the performance; and 2.0 means twice the performance of that reference profile. Component weights: 50% Redis RPS (pipeline=1, SET), 20% Redis RPS (pipeline=16, SET), 10% PassMark Memory Mark (composite), 10% Memory bandwidth (read, 16 MB ~ L3), 10% PassMark single-thread CPU. Rationale for component selection: In-memory key-value store workload, mixing direct Redis performance metrics with memory speed and latency benchmarks, and single-core CPU performance profiles.
Component Score Weight Impact
Raw Reference Normalized Target Actual
Redis RPS (pipeline=1, SET) 5,141,109 ops/sec 1,917,397 ops/sec 2.68 50.00% 50.00% +63.70%
Redis RPS (pipeline=16, SET) 34,564,529 ops/sec 11,681,410 ops/sec 2.96 20.00% 20.00% +24.20%
PassMark Memory Mark (composite) 2,917 2,418 1.21 10.00% 10.00% +1.92%
Memory bandwidth (read, 16 MB ~ L3) 119,992 MB/sec 108,140 MB/sec 1.11 10.00% 10.00% +1.05%
PassMark single-thread CPU 2,135 Mops/s 2,365 Mops/s 0.903 10.00% 10.00% -1.02%

Memory Bandwidth

Compression

OpenSSL

Geekbench Single-Core

Score: 1,332

Geekbench Multi-Core

Score: 14,069

Passmark CPU Scores

BENCHMARKSCORE
Mark
55615
Compression
953024
Encryption
68193
Extended Instructions
56952
Floating Point Maths
128866
Integer Maths
216068
Physics
5371
Prime Numbers
296
Single Threaded
2135
String Sorting
126195

Passmark Memory Scores

BENCHMARKSCORE
Memory Mark
2917
Database Operations
21347
Memory Latency
51
Memory Read Cached
24632
Memory Read Uncached
12768
Memory Write
13485

Stress-ng Raw Scores

Stress-ng Relative Multicore Performance

LLM Inference Speed for Prompt Processing

LLM Inference Speed for Text Generation

Static Web Server

Redis

Alternatives

Servers of the Same Family

INSTANCEvCPUsMEMORYGPUs
g5.xlarge 416 GiB1
g5.2xlarge 832 GiB1
g5.4xlarge 1664 GiB1
g5.8xlarge 32128 GiB1
g5.12xlarge 48192 GiB4
g5.24xlarge 96384 GiB4
g5.48xlarge 192768 GiB8

Similar Servers

INSTANCEVENDORvCPUsMEMORYGPUs

g5.16xlarge FAQs