logo

g5.8xlarge by Amazon Web Services

g5.8xlarge is a Graphics intensive Gen5 8xlarge server offered by Amazon Web Services with 32 vCPUs, 128 GiB of memory and 900 GB of storage. The pricing starts at 0.3769 USD per hour.
32 vCPU
128 GiB Memory
900 GB Storage
1 GPU
Spare SCore
12274
(All-cores)
766
(Single-core)

Specifications

Server Metadata

Vendor ID
aws
Name
g5.8xlarge
Description
Graphics intensive Gen5 8xlarge
Family
g5
Hw Virt
Status
active
Observed At
2026-08-04T10:53:37.461544

Availability

REGION / IDSPOTONDEMAND
Ohio (US) / us-east-2
0.951633 USD/h 2.448 USD/h
Oregon (US) / us-west-2
1.0204 USD/h 2.448 USD/h
Northern Virgina (US) / us-east-1
1.36012 USD/h 2.448 USD/h
Aragón (ES) / eu-south-2
- 2.5794 USD/h
Stockholm (SE) / eu-north-1
0.95425 USD/h 2.5964 USD/h
Quebec (CA) / ca-central-1
0.8783 USD/h 2.7181 USD/h
Dublin (IE) / eu-west-1
1.661367 USD/h 2.7327 USD/h
Tel Aviv (IL) / il-central-1
0.9224 USD/h 2.8693 USD/h
Mumbai (IN) / ap-south-1
1.1092 USD/h 2.9396 USD/h
United Arab Emirates / me-central-1
1.2293 USD/h 3.0028 USD/h
Seoul (KR) / ap-northeast-2
0.9547 USD/h 3.0101 USD/h
Frankfurt (DE) / eu-central-1
1.467833 USD/h 3.0612 USD/h
London (GB) / eu-west-2
0.8244 USD/h 3.1075 USD/h
Paris (FR) / eu-west-3
- 3.1075 USD/h
Sydney (AU) / ap-southeast-2
1.1403 USD/h 3.1829 USD/h
Jakarta (ID) / ap-southeast-3
0.91715 USD/h 3.4262 USD/h
Tokyo (JP) / ap-northeast-1
1.5978 USD/h 3.5503 USD/h
Hong Kong (HK) / ap-east-1
0.6856 USD/h 3.7689 USD/h
Sao Paulo (BR) / sa-east-1
2.2655 USD/h 4.1611 USD/h

Processor

vCPUs
32
Hypervisor
nitro
CPU Allocation
Dedicated
CPU Cores
16
CPU Speed
3.3 GHz
CPU Architecture
x86_64
CPU Manufacturer
AMD
CPU Family
EPYC
CPU Model
7R32
CPU L1D Cache
32 KiB
CPU L1D Cache Total
512 KiB
CPU L1I Cache
32 KiB
CPU L1I Cache Total
512 KiB
CPU L2 Cache
512 KiB
CPU L2 Cache Total
8 MiB
CPU L3 Cache
16 MiB
CPU L3 Cache Total
64 MiB
CPU Flags
fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, mmxext, fxsr_opt, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, nonstop_tsc, cpuid, extd_apicid, aperfmperf, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, sse4_1, sse4_2, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, cmp_legacy, cr8_legacy, abm, sse4a, misalignsse, 3dnowprefetch, topoext, ssbd, ibrs, ibpb, stibp, vmmcall, fsgsbase, bmi1, avx2, smep, bmi2, rdseed, adx, smap, clflushopt, clwb, sha_ni, xsaveopt, xsavec, xgetbv1, clzero, xsaveerptr, rdpru, wbnoinvd, arat, npt, nrip_save, rdpid
Ecpus
16
Scalability
100

System Resources and Accelerators

MEMORY
Memory Amount
128 GiB
Memory Amount Actual
128 GiB
Memory Generation
DDR4
Memory Speed
3200 Mhz
GPU
GPU Count
1
GPU Memory Min
22 GiB
GPU Memory Total
22 GiB
GPU Manufacturer
NVIDIA
GPU Family
Ampere
GPU Model
A10G
GPUs
  • NVIDIA Ampere NVIDIA A10G (Memory amount: 23028, Firmware version: 590.48.01, BIOS version: 94.02.91.00.07, Clock rate: 1710)
STORAGE
Storage Size
900 GB
Storage Type
nvme ssd
Storages
  • 900 GB nvme ssd
NETWORK
Network Speed Baseline
25 Gbps
Network Speed Max
25 Gbps
Network Storage Speed Baseline
16 Gbps
Network Storage Speed Max
16 Gbps
Inbound Traffic
0 GB/month
Outbound Traffic
0 GB/month
IPv4
0

CPU and System Topology

Server Description

An x86_64 GPU-accelerated instance featuring an NVIDIA A10G GPU, local NVMe storage, and top-tier performance for LLM inference workloads.

GPU AcceleratedGeneral PurposeStorage & Database

Amazon Web Services g5.8xlarge is an x86_64 graphics-intensive instance featuring 32 dedicated vCPUs from an AMD EPYC 7R32 processor running at 3.3 GHz, paired with 128.0 GB of DDR4 memory. It includes a single NVIDIA Ampere A10G GPU with 22 GB of VRAM and 900 GB of local NVMe SSD storage, supported by a 25 Gbps baseline network bandwidth. While single-core CPU benchmarks rank in the bottom 10%, the instance delivers top-tier performance in the top 10% for small and medium LLM inference tasks. Multi-core CPU and memory bandwidth benchmarks remain average. Qualitatively, the instance provides high resource density for GPU-dependent tasks, making it a viable option for machine learning inference, graphics rendering, and database operations that require local NVMe storage and hardware acceleration rather than raw single-threaded CPU performance.

Economics

Average Price per Region

Prices per Zone

Lowest Prices

Performance

Workload Profiles

1.23Score
Precomputed compound score for Cache Intensive workloads. A weighted average (geometric mean) of benchmark scores compared to their medians: score = ∏ (x_i / m_i)^(w_i / Σw). The score of 1.0 represents a synthetic baseline server with the median performance of each component benchmark; 0.5 means roughly half the performance; and 2.0 means twice the performance of that reference profile. Component weights: 50% Redis RPS (pipeline=1, SET), 20% Redis RPS (pipeline=16, SET), 10% PassMark Memory Mark (composite), 10% Memory bandwidth (read, 16 MB ~ L3), 10% PassMark single-thread CPU. Rationale for component selection: In-memory key-value store workload, mixing direct Redis performance metrics with memory speed and latency benchmarks, and single-core CPU performance profiles.
Component Score Weight Impact
Raw Reference Normalized Target Actual
Redis RPS (pipeline=1, SET) 2,617,513 ops/sec 1,910,660 ops/sec 1.37 50.00% 50.00% +17.00%
Redis RPS (pipeline=16, SET) 17,320,424 ops/sec 11,608,287 ops/sec 1.49 20.00% 20.00% +8.30%
PassMark Memory Mark (composite) 2,868 2,416 1.19 10.00% 10.00% +1.75%
Memory bandwidth (read, 16 MB ~ L3) 74,300 MB/sec 107,896 MB/sec 0.689 10.00% 10.00% -3.66%
PassMark single-thread CPU 2,136 Mops/s 2,364 Mops/s 0.903 10.00% 10.00% -1.02%

Memory Bandwidth

Compression

OpenSSL

Geekbench Single-Core

Score: 1,339

Geekbench Multi-Core

Score: 11,541

Passmark CPU Scores

BENCHMARKSCORE
Mark
34661
Compression
479822
Encryption
34138
Extended Instructions
29725
Floating Point Maths
64394
Integer Maths
107967
Physics
3311
Prime Numbers
162
Single Threaded
2136
String Sorting
66285

Passmark Memory Scores

BENCHMARKSCORE
Memory Mark
2868
Database Operations
11052
Memory Latency
50
Memory Read Cached
24687
Memory Read Uncached
12944
Memory Write
13346

Stress-ng Raw Scores

Stress-ng Relative Multicore Performance

LLM Inference Speed for Prompt Processing

LLM Inference Speed for Text Generation

Static Web Server

Redis

Alternatives

Servers of the Same Family

INSTANCEvCPUsMEMORYGPUs
g5.xlarge 416 GiB1
g5.2xlarge 832 GiB1
g5.4xlarge 1664 GiB1
g5.12xlarge 48192 GiB4
g5.16xlarge 64256 GiB1
g5.24xlarge 96384 GiB4
g5.48xlarge 192768 GiB8

Similar Servers

INSTANCEVENDORvCPUsMEMORYGPUs

g5.8xlarge FAQs