logo

gr6.8xlarge by Amazon Web Services

gr6.8xlarge is a Graphics intensive with a one to eight ratio of vCPU to memory Gen6 8xlarge server offered by Amazon Web Services with 32 vCPUs, 256 GiB of memory and 900 GB of storage. The pricing starts at 0.3883 USD per hour.
32 vCPU
256 GiB Memory
900 GB Storage
1 GPU
Spare SCore
47874
(All-cores)
2992
(Single-core)

Specifications

Server Metadata

Vendor ID
aws
Name
gr6.8xlarge
Description
Graphics intensive with a one to eight ratio of vCPU to memory Gen6 8xlarge
Family
gr6
Hw Virt
Status
active
Observed At
2026-07-31T18:14:00.877043

Availability

REGION / IDSPOTONDEMAND
Ohio (US) / us-east-2
0.705033 USD/h 2.4464 USD/h
Oregon (US) / us-west-2
0.8964 USD/h 2.4464 USD/h
Northern Virgina (US) / us-east-1
1.57465 USD/h 2.4464 USD/h
Aragón (ES) / eu-south-2
0.593633 USD/h 2.5777 USD/h
Stockholm (SE) / eu-north-1
0.6554 USD/h 2.5949 USD/h
Quebec (CA) / ca-central-1
0.53735 USD/h 2.7163 USD/h
Mumbai (IN) / ap-south-1
0.5722 USD/h 2.9377 USD/h
Seoul (KR) / ap-northeast-2
0.4464 USD/h 3.0081 USD/h
Frankfurt (DE) / eu-central-1
1.11965 USD/h 3.0593 USD/h
(MY) / ap-southeast-5
0.49485 USD/h 3.0815 USD/h
Paris (FR) / eu-west-3
0.59115 USD/h 3.1054 USD/h
London (GB) / eu-west-2
0.985 USD/h 3.1055 USD/h
Sydney (AU) / ap-southeast-2
1.26985 USD/h 3.1807 USD/h
Zurich (CH) / eu-central-2
0.7453 USD/h 3.4409 USD/h
Tokyo (JP) / ap-northeast-1
0.77605 USD/h 3.548 USD/h
Sao Paulo (BR) / sa-east-1
1.9951 USD/h 4.1585 USD/h

Processor

vCPUs
32
Hypervisor
nitro
CPU Allocation
Dedicated
CPU Cores
16
CPU Speed
3.4 GHz
CPU Architecture
x86_64
CPU Manufacturer
AMD
CPU Family
EPYC
CPU Model
7R13
CPU L1D Cache
32 KiB
CPU L1D Cache Total
512 KiB
CPU L1I Cache
32 KiB
CPU L1I Cache Total
512 KiB
CPU L2 Cache
512 KiB
CPU L2 Cache Total
8 MiB
CPU L3 Cache
32 MiB
CPU L3 Cache Total
64 MiB
CPU Flags
fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, mmxext, fxsr_opt, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, nonstop_tsc, cpuid, extd_apicid, aperfmperf, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, cmp_legacy, cr8_legacy, abm, sse4a, misalignsse, 3dnowprefetch, topoext, ssbd, ibrs, ibpb, stibp, vmmcall, fsgsbase, bmi1, avx2, smep, bmi2, invpcid, rdseed, adx, smap, clflushopt, clwb, sha_ni, xsaveopt, xsavec, xgetbv1, clzero, xsaveerptr, rdpru, wbnoinvd, arat, npt, nrip_save, vaes, vpclmulqdq, rdpid
Ecpus
16
Scalability
100

System Resources and Accelerators

MEMORY
Memory Amount
256 GiB
Memory Amount Actual
256 GiB
Memory Generation
DDR4
Memory Speed
3200 Mhz
GPU
GPU Count
1
GPU Memory Min
22 GiB
GPU Memory Total
22 GiB
GPU Manufacturer
NVIDIA
GPU Family
Ada Lovelace
GPU Model
L4
GPUs
  • NVIDIA Ada Lovelace NVIDIA L4 (Memory amount: 23034, Firmware version: 590.48.01, BIOS version: 95.04.65.00.37, Clock rate: 2040)
STORAGE
Storage Size
900 GB
Storage Type
nvme ssd
Storages
  • 450 GB nvme ssd
  • 450 GB nvme ssd
NETWORK
Network Speed Baseline
25 Gbps
Network Speed Max
25 Gbps
Network Storage Speed Baseline
16 Gbps
Network Storage Speed Max
16 Gbps
Inbound Traffic
0 GB/month
Outbound Traffic
0 GB/month
IPv4
0

CPU and System Topology

Server Description

A GPU-accelerated, memory-optimized instance featuring an NVIDIA L4 GPU and local NVMe storage for high-throughput machine learning inference.

GPU AcceleratedMemory OptimizedStorage & Database

Amazon Web Services gr6.8xlarge is a graphics-intensive, memory-optimized instance featuring 32 vCPUs and 16 physical cores powered by an AMD EPYC 7R13 processor at 3.4 GHz. It is equipped with 256 GB of DDR4 memory, a single NVIDIA L4 GPU with 22 GB of VRAM, and 900 GB of local NVMe SSD storage. Operating on the Nitro hypervisor, the instance delivers 25 Gbps baseline network bandwidth. Benchmark data highlights top-tier performance in LLM inference for small and medium models, alongside strong single-core CPU and memory performance. While multi-core CPU benchmarks remain average, the server's high memory-to-vCPU ratio and dedicated GPU acceleration make it highly efficient for machine learning inference, image processing, and graphics-heavy workloads.

Economics

Average Price per Region

Prices per Zone

Lowest Prices

Performance

Workload Profiles

1.34Score
Precomputed compound score for Cache Intensive workloads. A weighted average (geometric mean) of benchmark scores compared to their medians: score = ∏ (x_i / m_i)^(w_i / Σw). The score of 1.0 represents a synthetic baseline server with the median performance of each component benchmark; 0.5 means roughly half the performance; and 2.0 means twice the performance of that reference profile. Component weights: 50% Redis RPS (pipeline=1, SET), 20% Redis RPS (pipeline=16, SET), 10% PassMark Memory Mark (composite), 10% Memory bandwidth (read, 16 MB ~ L3), 10% PassMark single-thread CPU. Rationale for component selection: In-memory key-value store workload, mixing direct Redis performance metrics with memory speed and latency benchmarks, and single-core CPU performance profiles.
Component Score Weight Impact
Raw Reference Normalized Target Actual
Redis RPS (pipeline=1, SET) 2,773,170 ops/sec 1,896,415 ops/sec 1.46 50.00% 50.00% +20.80%
Redis RPS (pipeline=16, SET) 19,962,308 ops/sec 11,539,523 ops/sec 1.73 20.00% 20.00% +11.60%
PassMark Memory Mark (composite) 2,795 2,411 1.16 10.00% 10.00% +1.50%
Memory bandwidth (read, 16 MB ~ L3) 78,848 MB/sec 107,368 MB/sec 0.734 10.00% 10.00% -3.05%
PassMark single-thread CPU 2,639 Mops/s 2,363 Mops/s 1.12 10.00% 10.00% +1.14%

Memory Bandwidth

Compression

OpenSSL

Geekbench Single-Core

Score: 1,733

Geekbench Multi-Core

Score: 14,220

Passmark CPU Scores

BENCHMARKSCORE
Mark
41924
Compression
560028
Encryption
37916
Extended Instructions
33099
Floating Point Maths
80715
Integer Maths
147220
Physics
3982
Prime Numbers
198
Single Threaded
2639
String Sorting
75772

Passmark Memory Scores

BENCHMARKSCORE
Memory Mark
2795
Database Operations
14066
Memory Latency
65
Memory Read Cached
26624
Memory Read Uncached
17307
Memory Write
17608

Stress-ng Raw Scores

Stress-ng Relative Multicore Performance

LLM Inference Speed for Prompt Processing

LLM Inference Speed for Text Generation

Static Web Server

Redis

Alternatives

Servers of the Same Family

INSTANCEvCPUsMEMORYGPUs
gr6.4xlarge 16128 GiB1

Similar Servers

INSTANCEVENDORvCPUsMEMORYGPUs

gr6.8xlarge FAQs