logo

g6.16xlarge by Amazon Web Services

g6.16xlarge is a Graphics intensive Gen6 16xlarge server offered by Amazon Web Services with 64 vCPUs, 256 GiB of memory and 1.88 TB of storage. The pricing starts at 0.7538 USD per hour.
64 vCPU
256 GiB Memory
1.88 TB Storage
1 GPU
Spare SCore
95679
(All-cores)
2993
(Single-core)

Specifications

Server Metadata

Vendor ID
aws
Name
g6.16xlarge
Description
Graphics intensive Gen6 16xlarge
Family
g6
Hw Virt
Status
active
Observed At
2026-07-21T16:53:04.515109

Availability

REGION / IDSPOTONDEMAND
Ohio (US) / us-east-2
1.280367 USD/h 3.3968 USD/h
Oregon (US) / us-west-2
1.71255 USD/h 3.3968 USD/h
Northern Virgina (US) / us-east-1
1.76294 USD/h 3.3968 USD/h
Aragón (ES) / eu-south-2
0.889 USD/h 3.5791 USD/h
Stockholm (SE) / eu-north-1
1.09625 USD/h 3.6028 USD/h
Quebec (CA) / ca-central-1
1.5609 USD/h 3.7716 USD/h
Mumbai (IN) / ap-south-1
0.9234 USD/h 4.0789 USD/h
United Arab Emirates / me-central-1
1.7075 USD/h 4.171 USD/h
Seoul (KR) / ap-northeast-2
1.0722 USD/h 4.1768 USD/h
Frankfurt (DE) / eu-central-1
1.495467 USD/h 4.2477 USD/h
(MY) / ap-southeast-5
0.894 USD/h 4.2789 USD/h
Paris (FR) / eu-west-3
1.39135 USD/h 4.3118 USD/h
London (GB) / eu-west-2
1.475 USD/h 4.3118 USD/h
Sydney (AU) / ap-southeast-2
1.8283 USD/h 4.4165 USD/h
Zurich (CH) / eu-central-2
1.15475 USD/h 4.5794 USD/h
Tokyo (JP) / ap-northeast-1
1.8347 USD/h 4.9264 USD/h
Sao Paulo (BR) / sa-east-1
2.69195 USD/h 5.7739 USD/h

Processor

vCPUs
64
Hypervisor
nitro
CPU Allocation
Dedicated
CPU Cores
32
CPU Speed
3.4 GHz
CPU Architecture
x86_64
CPU Manufacturer
AMD
CPU Family
EPYC
CPU Model
7R13
CPU L1D Cache
32 KiB
CPU L1D Cache Total
1 MiB
CPU L1I Cache
32 KiB
CPU L1I Cache Total
1 MiB
CPU L2 Cache
512 KiB
CPU L2 Cache Total
16 MiB
CPU L3 Cache
32 MiB
CPU L3 Cache Total
128 MiB
CPU Flags
fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, mmxext, fxsr_opt, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, nonstop_tsc, cpuid, extd_apicid, aperfmperf, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, cmp_legacy, cr8_legacy, abm, sse4a, misalignsse, 3dnowprefetch, topoext, ssbd, ibrs, ibpb, stibp, vmmcall, fsgsbase, bmi1, avx2, smep, bmi2, invpcid, rdseed, adx, smap, clflushopt, clwb, sha_ni, xsaveopt, xsavec, xgetbv1, clzero, xsaveerptr, rdpru, wbnoinvd, arat, npt, nrip_save, vaes, vpclmulqdq, rdpid
Ecpus
32
Scalability
100

System Resources and Accelerators

MEMORY
Memory Amount
256 GiB
Memory Amount Actual
256 GiB
Memory Generation
DDR4
Memory Speed
3200 Mhz
GPU
GPU Count
1
GPU Memory Min
22 GiB
GPU Memory Total
22 GiB
GPU Manufacturer
NVIDIA
GPU Family
Ada Lovelace
GPU Model
L4
GPUs
  • NVIDIA Ada Lovelace NVIDIA L4 (Memory amount: 23034, Firmware version: 590.48.01, BIOS version: 95.04.65.00.37, Clock rate: 2040)
STORAGE
Storage Size
1880 GB
Storage Type
nvme ssd
Storages
  • 940 GB nvme ssd
  • 940 GB nvme ssd
NETWORK
Network Speed Baseline
25 Gbps
Network Speed Max
25 Gbps
Network Storage Speed Baseline
20 Gbps
Network Storage Speed Max
20 Gbps
Inbound Traffic
0 GB/month
Outbound Traffic
0 GB/month
IPv4
0

CPU and System Topology

Server Description

A GPU-accelerated virtual machine featuring dedicated AMD EPYC processors, local NVMe storage, and balanced performance for graphics and machine learning workloads.

GPU AcceleratedCompute OptimizedStorage & Database

Amazon Web Services g6.16xlarge is a graphics-intensive instance featuring 64 vCPUs from an AMD EPYC 7R13 processor, 256.0 GB of DDR4 memory, and a single NVIDIA L4 GPU with 22 GB of VRAM. Built on the Nitro hypervisor, it includes 1880 GB of local NVMe SSD storage and 25 Gbps baseline network bandwidth. Benchmarks show strong performance in multi-core CPU tasks, ray tracing, and database operations. In LLM inference, the instance achieves top-tier results for small and medium models, though text generation for large 70B models drops to the bottom 25% due to single-GPU constraints. This server is well-suited for graphics rendering, image processing, and medium-scale machine learning workloads.

Economics

Average Price per Region

Prices per Zone

Lowest Prices

Performance

Workload Profiles

2.17Score
Precomputed compound score for Cache Intensive workloads. A weighted average (geometric mean) of benchmark scores compared to their medians: score = ∏ (x_i / m_i)^(w_i / Σw). The score of 1.0 represents a synthetic baseline server with the median performance of each component benchmark; 0.5 means roughly half the performance; and 2.0 means twice the performance of that reference profile. Component weights: 50% Redis RPS (pipeline=1, SET), 20% Redis RPS (pipeline=16, SET), 10% PassMark Memory Mark (composite), 10% Memory bandwidth (read, 16 MB ~ L3), 10% PassMark single-thread CPU. Rationale for component selection: In-memory key-value store workload, mixing direct Redis performance metrics with memory speed and latency benchmarks, and single-core CPU performance profiles.
Component Score Weight Impact
Raw Reference Normalized Target Actual
Redis RPS (pipeline=1, SET) 5,236,957 ops/sec 1,896,415 ops/sec 2.76 50.00% 50.00% +66.10%
Redis RPS (pipeline=16, SET) 36,549,674 ops/sec 11,539,523 ops/sec 3.17 20.00% 20.00% +26.00%
PassMark Memory Mark (composite) 2,777 2,411 1.15 10.00% 10.00% +1.41%
Memory bandwidth (read, 16 MB ~ L3) 120,906 MB/sec 107,368 MB/sec 1.13 10.00% 10.00% +1.23%
PassMark single-thread CPU 2,640 Mops/s 2,363 Mops/s 1.12 10.00% 10.00% +1.14%

Memory Bandwidth

Compression

OpenSSL

Geekbench Single-Core

Score: 1,736

Geekbench Multi-Core

Score: 16,623

Passmark CPU Scores

BENCHMARKSCORE
Mark
60670
Compression
897370
Encryption
65193
Extended Instructions
47745
Floating Point Maths
144669
Integer Maths
269588
Physics
6174
Prime Numbers
341
Single Threaded
2640
String Sorting
119896

Passmark Memory Scores

BENCHMARKSCORE
Memory Mark
2777
Database Operations
24127
Memory Latency
67
Memory Read Cached
26410
Memory Read Uncached
16540
Memory Write
17058

Stress-ng Raw Scores

Stress-ng Relative Multicore Performance

LLM Inference Speed for Prompt Processing

LLM Inference Speed for Text Generation

Static Web Server

Redis

Alternatives

Servers of the Same Family

INSTANCEvCPUsMEMORYGPUs
g6.xlarge 416 GiB1
g6.2xlarge 832 GiB1
g6.4xlarge 1664 GiB1
g6.8xlarge 32128 GiB1
g6.12xlarge 48192 GiB4
g6.24xlarge 96384 GiB4
g6.48xlarge 192768 GiB8

Similar Servers

INSTANCEVENDORvCPUsMEMORYGPUs

g6.16xlarge FAQs