g5.16xlarge by Amazon Web Services
(All-cores)
(Single-core)
Specifications
Server Metadata
Vendor ID | aws |
Name | g5.16xlarge |
Description | Graphics intensive Gen5 16xlarge |
Family | g5 |
Hw Virt | |
Status | active |
Observed At | 2026-07-25T19:53:22.718403 |
Availability
| REGION / ID | SPOT | ONDEMAND |
|---|---|---|
Ohio (US) / us-east-2 | 1.109533 USD/h | 4.096 USD/h |
Oregon (US) / us-west-2 | 1.647667 USD/h | 4.096 USD/h |
Northern Virgina (US) / us-east-1 | 2.09186 USD/h | 4.096 USD/h |
Stockholm (SE) / eu-north-1 | 1.14505 USD/h | 4.3444 USD/h |
Quebec (CA) / ca-central-1 | 1.09075 USD/h | 4.5479 USD/h |
Dublin (IE) / eu-west-1 | 1.2592 USD/h | 4.5724 USD/h |
Tel Aviv (IL) / il-central-1 | 2.443 USD/h | 4.801 USD/h |
Mumbai (IN) / ap-south-1 | 1.3648 USD/h | 4.9185 USD/h |
United Arab Emirates / me-central-1 | 2.0568 USD/h | 5.0243 USD/h |
Seoul (KR) / ap-northeast-2 | 1.260033 USD/h | 5.0365 USD/h |
Frankfurt (DE) / eu-central-1 | 2.0036 USD/h | 5.122 USD/h |
London (GB) / eu-west-2 | 2.28895 USD/h | 5.1994 USD/h |
Sydney (AU) / ap-southeast-2 | 1.53215 USD/h | 5.3256 USD/h |
Jakarta (ID) / ap-southeast-3 | 3.0284 USD/h | 5.7328 USD/h |
Tokyo (JP) / ap-northeast-1 | 1.1138 USD/h | 5.9404 USD/h |
Hong Kong (HK) / ap-east-1 | 2.03135 USD/h | 6.306 USD/h |
Sao Paulo (BR) / sa-east-1 | 5.7176 USD/h | 6.9624 USD/h |
Processor
vCPUs | 64 |
Hypervisor | nitro |
CPU Allocation | Dedicated |
CPU Cores | 32 |
CPU Speed | 3.3 GHz |
CPU Architecture | x86_64 |
CPU Manufacturer | AMD |
CPU Family | EPYC |
CPU Model | 7R32 |
CPU L1D Cache | 32 KiB |
CPU L1D Cache Total | 1 MiB |
CPU L1I Cache | 32 KiB |
CPU L1I Cache Total | 1 MiB |
CPU L2 Cache | 512 KiB |
CPU L2 Cache Total | 16 MiB |
CPU L3 Cache | 16 MiB |
CPU L3 Cache Total | 128 MiB |
CPU Flags | fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, mmxext, fxsr_opt, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, nonstop_tsc, cpuid, extd_apicid, aperfmperf, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, sse4_1, sse4_2, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, cmp_legacy, cr8_legacy, abm, sse4a, misalignsse, 3dnowprefetch, topoext, ssbd, ibrs, ibpb, stibp, vmmcall, fsgsbase, bmi1, avx2, smep, bmi2, rdseed, adx, smap, clflushopt, clwb, sha_ni, xsaveopt, xsavec, xgetbv1, clzero, xsaveerptr, rdpru, wbnoinvd, arat, npt, nrip_save, rdpid |
Ecpus | 32.1 |
Scalability | 100.31 |
System Resources and Accelerators
| MEMORY | |
|---|---|
Memory Amount | 256 GiB |
Memory Amount Actual | 256 GiB |
Memory Generation | DDR4 |
Memory Speed | 3200 Mhz |
| GPU | |
|---|---|
GPU Count | 1 |
GPU Memory Min | 22 GiB |
GPU Memory Total | 22 GiB |
GPU Manufacturer | NVIDIA |
GPU Family | Ampere |
GPU Model | A10G |
GPUs |
|
| STORAGE | |
|---|---|
Storage Size | 1900 GB |
Storage Type | nvme ssd |
Storages |
|
| NETWORK | |
|---|---|
Network Speed Baseline | 25 Gbps |
Network Speed Max | 25 Gbps |
Network Storage Speed Baseline | 16 Gbps |
Network Storage Speed Max | 16 Gbps |
Inbound Traffic | 0 GB/month |
Outbound Traffic | 0 GB/month |
IPv4 | 0 |
CPU and System Topology
Server Description
An accelerated virtual server combining dedicated AMD EPYC processors, local NVMe storage, and NVIDIA Ampere graphics for demanding machine learning workloads.
Amazon Web Services g5.16xlarge is a graphics-intensive instance featuring 64 dedicated vCPUs from an AMD EPYC 7R32 processor running at 3.3 GHz, 256.0 GB of DDR4 memory, and a baseline network bandwidth of 25 Gbps. It includes a bundled NVIDIA Ampere A10G GPU with 22 GB of VRAM and 1900 GB of local NVMe SSD storage. Performance benchmarks show a contrast between poor single-core CPU metrics and top-tier LLM inference speeds for small and medium models. While PassMark CPU and memory scores are strong, text generation for large 70B models is weak. This configuration offers qualitative cost efficiency by combining dedicated compute, high-speed local storage, and GPU acceleration on the Nitro hypervisor. It is best suited for machine learning inference, ray tracing, and workloads requiring high-density GPU and storage resources.
Economics
Performance
Memory Bandwidth
Compression
OpenSSL
Geekbench Single-Core
Geekbench Multi-Core
Passmark CPU Scores
| BENCHMARK | SCORE |
|---|---|
Mark | 55615 |
Compression | 953024 |
Encryption | 68193 |
Extended Instructions | 56952 |
Floating Point Maths | 128866 |
Integer Maths | 216068 |
Physics | 5371 |
Prime Numbers | 296 |
Single Threaded | 2135 |
String Sorting | 126195 |
Passmark Memory Scores
| BENCHMARK | SCORE |
|---|---|
Memory Mark | 2917 |
Database Operations | 21347 |
Memory Latency | 51 |
Memory Read Cached | 24632 |
Memory Read Uncached | 12768 |
Memory Write | 13485 |
Stress-ng Raw Scores
Stress-ng Relative Multicore Performance
LLM Inference Speed for Prompt Processing
LLM Inference Speed for Text Generation
Static Web Server
Redis
Alternatives
Servers of the Same Family
| INSTANCE | vCPUs | MEMORY | GPUs |
|---|---|---|---|
| g5.xlarge | 4 | 16 GiB | 1 |
| g5.2xlarge | 8 | 32 GiB | 1 |
| g5.4xlarge | 16 | 64 GiB | 1 |
| g5.8xlarge | 32 | 128 GiB | 1 |
| g5.12xlarge | 48 | 192 GiB | 4 |
| g5.24xlarge | 96 | 384 GiB | 4 |
| g5.48xlarge | 192 | 768 GiB | 8 |
Similar Servers
| INSTANCE | VENDOR | vCPUs | MEMORY | GPUs |
|---|