logo

g5.4xlarge by Amazon Web Services

g5.4xlarge is a Graphics intensive Gen5 4xlarge server offered by Amazon Web Services with 16 vCPUs, 64 GiB of memory and 600 GB of storage. The pricing starts at 0.442 USD per hour.
16 vCPU
64 GiB Memory
600 GB Storage
1 GPU
Spare SCore
6147
(All-cores)
767
(Single-core)

Specifications

Server Metadata

Vendor ID
aws
Name
g5.4xlarge
Description
Graphics intensive Gen5 4xlarge
Family
g5
Hw Virt
Status
active
Observed At
2026-07-21T01:33:20.913921

Availability

REGION / IDSPOTONDEMAND
Oregon (US) / us-west-2
0.637 USD/h 1.624 USD/h
Ohio (US) / us-east-2
0.788133 USD/h 1.624 USD/h
Northern Virgina (US) / us-east-1
0.95948 USD/h 1.624 USD/h
Stockholm (SE) / eu-north-1
0.533 USD/h 1.7225 USD/h
Quebec (CA) / ca-central-1
0.9283 USD/h 1.8032 USD/h
Dublin (IE) / eu-west-1
1.104033 USD/h 1.8129 USD/h
Tel Aviv (IL) / il-central-1
1.4881 USD/h 1.9035 USD/h
Mumbai (IN) / ap-south-1
0.5717 USD/h 1.9501 USD/h
United Arab Emirates / me-central-1
0.8155 USD/h 1.9921 USD/h
Seoul (KR) / ap-northeast-2
0.764367 USD/h 1.9969 USD/h
Frankfurt (DE) / eu-central-1
1.0618 USD/h 2.0308 USD/h
London (GB) / eu-west-2
0.77425 USD/h 2.0615 USD/h
Sydney (AU) / ap-southeast-2
1.0361 USD/h 2.1115 USD/h
Jakarta (ID) / ap-southeast-3
0.452 USD/h 2.2729 USD/h
Tokyo (JP) / ap-northeast-1
1.00985 USD/h 2.3553 USD/h
Hong Kong (HK) / ap-east-1
1.2939 USD/h 2.5002 USD/h
Sao Paulo (BR) / sa-east-1
2.2528 USD/h 2.7605 USD/h

Processor

vCPUs
16
Hypervisor
nitro
CPU Allocation
Dedicated
CPU Cores
8
CPU Speed
3.3 GHz
CPU Architecture
x86_64
CPU Manufacturer
AMD
CPU Family
EPYC
CPU Model
7R32
CPU L1D Cache
32 KiB
CPU L1D Cache Total
256 KiB
CPU L1I Cache
32 KiB
CPU L1I Cache Total
256 KiB
CPU L2 Cache
512 KiB
CPU L2 Cache Total
4 MiB
CPU L3 Cache
16 MiB
CPU L3 Cache Total
32 MiB
CPU Flags
fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, mmxext, fxsr_opt, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, nonstop_tsc, cpuid, extd_apicid, aperfmperf, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, sse4_1, sse4_2, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, cmp_legacy, cr8_legacy, abm, sse4a, misalignsse, 3dnowprefetch, topoext, ssbd, ibrs, ibpb, stibp, vmmcall, fsgsbase, bmi1, avx2, smep, bmi2, rdseed, adx, smap, clflushopt, clwb, sha_ni, xsaveopt, xsavec, xgetbv1, clzero, xsaveerptr, rdpru, wbnoinvd, arat, npt, nrip_save, rdpid
Ecpus
8
Scalability
100

System Resources and Accelerators

MEMORY
Memory Amount
64 GiB
Memory Amount Actual
64 GiB
Memory Generation
DDR4
Memory Speed
3200 Mhz
GPU
GPU Count
1
GPU Memory Min
22 GiB
GPU Memory Total
22 GiB
GPU Manufacturer
NVIDIA
GPU Family
Ampere
GPU Model
A10G
GPUs
  • NVIDIA Ampere NVIDIA A10G (Memory amount: 23028, Firmware version: 590.48.01, BIOS version: 94.02.75.00.01, Clock rate: 1710)
STORAGE
Storage Size
600 GB
Storage Type
nvme ssd
Storages
  • 600 GB nvme ssd
NETWORK
Network Speed Baseline
10 Gbps
Network Speed Max
25 Gbps
Network Storage Speed Baseline
4.75 Gbps
Network Storage Speed Max
4.75 Gbps
Inbound Traffic
0 GB/month
Outbound Traffic
0 GB/month
IPv4
0

CPU and System Topology

Server Description

A GPU-accelerated instance featuring dedicated NVMe storage and an NVIDIA Ampere GPU, optimized for machine learning inference and graphics-intensive workloads.

GPU AcceleratedStorage & DatabaseGeneral Purpose

Amazon Web Services g5.4xlarge is a GPU-accelerated server powered by the Nitro hypervisor. It features an AMD EPYC 7R32 processor with 16 vCPUs, 64 GB of DDR4 memory, 600 GB of local NVMe SSD storage, and a single NVIDIA Ampere A10G GPU with 22 GB of VRAM. While synthetic CPU and memory bandwidth benchmarks place the instance in the average to weak categories, its machine learning capabilities are highly competitive. Specifically, the server achieves top-tier performance in LLM inference for small and medium models, and strong performance for large model prompt processing. This hardware profile makes the g5.4xlarge highly cost-efficient for GPU-driven workloads, including low-to-medium parameter LLM deployment, ray tracing, and graphics rendering, whereas it presents tradeoffs for pure CPU-bound or high-memory-bandwidth applications.

Economics

Average Price per Region

Prices per Zone

Lowest Prices

Performance

Workload Profiles

0.707Score
Precomputed compound score for Cache Intensive workloads. A weighted average (geometric mean) of benchmark scores compared to their medians: score = ∏ (x_i / m_i)^(w_i / Σw). The score of 1.0 represents a synthetic baseline server with the median performance of each component benchmark; 0.5 means roughly half the performance; and 2.0 means twice the performance of that reference profile. Component weights: 50% Redis RPS (pipeline=1, SET), 20% Redis RPS (pipeline=16, SET), 10% PassMark Memory Mark (composite), 10% Memory bandwidth (read, 16 MB ~ L3), 10% PassMark single-thread CPU. Rationale for component selection: In-memory key-value store workload, mixing direct Redis performance metrics with memory speed and latency benchmarks, and single-core CPU performance profiles.
Component Score Weight Impact
Raw Reference Normalized Target Actual
Redis RPS (pipeline=1, SET) 1,311,632 ops/sec 1,896,415 ops/sec 0.692 50.00% 50.00% -16.80%
Redis RPS (pipeline=16, SET) 8,635,272 ops/sec 11,539,523 ops/sec 0.748 20.00% 20.00% -5.64%
PassMark Memory Mark (composite) 2,668 2,411 1.11 10.00% 10.00% +1.05%
Memory bandwidth (read, 16 MB ~ L3) 37,793 MB/sec 107,368 MB/sec 0.352 10.00% 10.00% -9.91%
PassMark single-thread CPU 2,133 Mops/s 2,363 Mops/s 0.902 10.00% 10.00% -1.03%

Memory Bandwidth

Compression

OpenSSL

Geekbench Single-Core

Score: 1,340

Geekbench Multi-Core

Score: 8,024

Passmark CPU Scores

BENCHMARKSCORE
Mark
19545
Compression
237423
Encryption
17037
Extended Instructions
14785
Floating Point Maths
32192
Integer Maths
53922
Physics
1750
Prime Numbers
84
Single Threaded
2133
String Sorting
33351

Passmark Memory Scores

BENCHMARKSCORE
Memory Mark
2668
Database Operations
5612
Memory Latency
51
Memory Read Cached
24688
Memory Read Uncached
12822
Memory Write
13221

Stress-ng Raw Scores

Stress-ng Relative Multicore Performance

LLM Inference Speed for Prompt Processing

LLM Inference Speed for Text Generation

Static Web Server

Redis

Alternatives

Servers of the Same Family

INSTANCEvCPUsMEMORYGPUs
g5.xlarge 416 GiB1
g5.2xlarge 832 GiB1
g5.8xlarge 32128 GiB1
g5.12xlarge 48192 GiB4
g5.16xlarge 64256 GiB1
g5.24xlarge 96384 GiB4
g5.48xlarge 192768 GiB8

Similar Servers

INSTANCEVENDORvCPUsMEMORYGPUs

g5.4xlarge FAQs