g5.16xlarge by Amazon Web Services
(All-cores)
(Single-core)
Specifications
Server Metadata
Vendor ID | aws |
Name | g5.16xlarge |
Description | Graphics intensive Gen5 16xlarge |
Family | g5 |
Hw Virt | |
Average Time To Start | 11 |
Status | active |
Observed At | 2026-09-08T20:54:13.156789 |
Availability
| REGION / ID | SPOT | ONDEMAND |
|---|---|---|
Oregon (US) / us-west-2 | 1.303033 USD/h | 4.096 USD/h |
Ohio (US) / us-east-2 | 1.9862 USD/h | 4.096 USD/h |
Northern Virgina (US) / us-east-1 | 2.4364 USD/h | 4.096 USD/h |
Aragón (ES) / eu-south-2 | 4.12365 USD/h | 4.3159 USD/h |
Stockholm (SE) / eu-north-1 | 0.72085 USD/h | 4.3444 USD/h |
Quebec (CA) / ca-central-1 | 1.6867 USD/h | 4.5479 USD/h |
Dublin (IE) / eu-west-1 | 1.5766 USD/h | 4.5724 USD/h |
Tel Aviv (IL) / il-central-1 | 1.945033 USD/h | 4.801 USD/h |
Mumbai (IN) / ap-south-1 | 1.73485 USD/h | 4.9185 USD/h |
United Arab Emirates / me-central-1 | 2.0568 USD/h | 5.0243 USD/h |
Seoul (KR) / ap-northeast-2 | 2.343467 USD/h | 5.0365 USD/h |
Frankfurt (DE) / eu-central-1 | 1.9861 USD/h | 5.122 USD/h |
Paris (FR) / eu-west-3 | 2.69165 USD/h | 5.1994 USD/h |
London (GB) / eu-west-2 | 3.39245 USD/h | 5.1994 USD/h |
Sydney (AU) / ap-southeast-2 | 1.32 USD/h | 5.3256 USD/h |
Jakarta (ID) / ap-southeast-3 | 2.8133 USD/h | 5.7328 USD/h |
Tokyo (JP) / ap-northeast-1 | 2.7046 USD/h | 5.9404 USD/h |
Hong Kong (HK) / ap-east-1 | 1.4415 USD/h | 6.306 USD/h |
Sao Paulo (BR) / sa-east-1 | 5.9081 USD/h | 6.9624 USD/h |
Processor
vCPUs | 64 |
Hypervisor | nitro |
CPU Allocation | Dedicated |
CPU Cores | 32 |
CPU Speed | 3.3 GHz |
CPU Architecture | x86_64 |
CPU Manufacturer | AMD |
CPU Family | EPYC |
CPU Model | 7R32 |
CPU L1D Cache | 32 KiB |
CPU L1D Cache Total | 1 MiB |
CPU L1I Cache | 32 KiB |
CPU L1I Cache Total | 1 MiB |
CPU L2 Cache | 512 KiB |
CPU L2 Cache Total | 16 MiB |
CPU L3 Cache | 16 MiB |
CPU L3 Cache Total | 128 MiB |
CPU Flags | fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, mmxext, fxsr_opt, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, nonstop_tsc, cpuid, extd_apicid, aperfmperf, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, sse4_1, sse4_2, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, cmp_legacy, cr8_legacy, abm, sse4a, misalignsse, 3dnowprefetch, topoext, ssbd, ibrs, ibpb, stibp, vmmcall, fsgsbase, bmi1, avx2, smep, bmi2, rdseed, adx, smap, clflushopt, clwb, sha_ni, xsaveopt, xsavec, xgetbv1, clzero, xsaveerptr, rdpru, wbnoinvd, arat, npt, nrip_save, rdpid |
Ecpus | 32.1 |
Scalability | 100.31 |
System Resources and Accelerators
| MEMORY | |
|---|---|
Memory Amount | 256 GiB |
Memory Amount Actual | 256 GiB |
Memory Generation | DDR4 |
Memory Speed | 3200 Mhz |
| GPU | |
|---|---|
GPU Count | 1 |
GPU Memory Min | 22 GiB |
GPU Memory Total | 22 GiB |
GPU Manufacturer | NVIDIA |
GPU Family | Ampere |
GPU Model | A10G |
GPUs |
|
| STORAGE | |
|---|---|
Storage Size | 1900 GB |
Storage Type | nvme ssd |
Storages |
|
| NETWORK | |
|---|---|
Network Speed Baseline | 25 Gbps |
Network Speed Max | 25 Gbps |
Network Storage Speed Baseline | 16 Gbps |
Network Storage Speed Max | 16 Gbps |
Inbound Traffic | 0 GB/month |
Outbound Traffic | 0 GB/month |
IPv4 | 0 |
CPU and System Topology
Server Description
An accelerated virtual server combining dedicated AMD EPYC processors, local NVMe storage, and NVIDIA Ampere graphics for demanding machine learning workloads.
Amazon Web Services g5.16xlarge is a graphics-intensive instance featuring 64 dedicated vCPUs from an AMD EPYC 7R32 processor running at 3.3 GHz, 256.0 GB of DDR4 memory, and a baseline network bandwidth of 25 Gbps. It includes a bundled NVIDIA Ampere A10G GPU with 22 GB of VRAM and 1900 GB of local NVMe SSD storage. Performance benchmarks show a contrast between poor single-core CPU metrics and top-tier LLM inference speeds for small and medium models. While PassMark CPU and memory scores are strong, text generation for large 70B models is weak. This configuration offers qualitative cost efficiency by combining dedicated compute, high-speed local storage, and GPU acceleration on the Nitro hypervisor. It is best suited for machine learning inference, ray tracing, and workloads requiring high-density GPU and storage resources.
Economics
Performance
Memory Bandwidth
Compression
OpenSSL
Geekbench Single-Core
Geekbench Multi-Core
Passmark CPU Scores
| BENCHMARK | SCORE |
|---|---|
Mark | 55615 |
Compression | 953024 |
Encryption | 68193 |
Extended Instructions | 56952 |
Floating Point Maths | 128866 |
Integer Maths | 216068 |
Physics | 5371 |
Prime Numbers | 296 |
Single Threaded | 2135 |
String Sorting | 126195 |
Passmark Memory Scores
| BENCHMARK | SCORE |
|---|---|
Memory Mark | 2917 |
Database Operations | 21347 |
Memory Latency | 51 |
Memory Read Cached | 24632 |
Memory Read Uncached | 12768 |
Memory Write | 13485 |
Stress-ng Raw Scores
Stress-ng Relative Multicore Performance
LLM Inference Speed for Prompt Processing
LLM Inference Speed for Text Generation
Static Web Server
Redis
Alternatives
Servers of the Same Family
| INSTANCE | vCPUs | MEMORY | GPUs |
|---|---|---|---|
| g5.xlarge | 4 | 16 GiB | 1 |
| g5.2xlarge | 8 | 32 GiB | 1 |
| g5.4xlarge | 16 | 64 GiB | 1 |
| g5.8xlarge | 32 | 128 GiB | 1 |
| g5.12xlarge | 48 | 192 GiB | 4 |
| g5.24xlarge | 96 | 384 GiB | 4 |
| g5.48xlarge | 192 | 768 GiB | 8 |
Similar Servers
| INSTANCE | VENDOR | vCPUs | MEMORY | GPUs |
|---|