logo

inf2.48xlarge by Amazon Web Services

inf2.48xlarge is a AWS Inferentia Gen2 48xlarge server offered by Amazon Web Services with 192 vCPUs, 768 GiB of memory and 0 GB of storage. The pricing starts at 1.2981 USD per hour.
192 vCPU
768 GiB Memory
Spare SCore
286883
(All-cores)
2992
(Single-core)

Specifications

Server Metadata

Vendor ID
aws
Name
inf2.48xlarge
Description
AWS Inferentia Gen2 48xlarge
Family
inf2
Hw Virt
Status
active
Observed At
2026-07-31T20:54:49.728062

Availability

REGION / IDSPOTONDEMAND
Ohio (US) / us-east-2
1.5397 USD/h 12.9813 USD/h
Oregon (US) / us-west-2
2.678225 USD/h 12.9813 USD/h
Northern Virgina (US) / us-east-1
5.78595 USD/h 12.9813 USD/h
Stockholm (SE) / eu-north-1
1.8757 USD/h 14.2794 USD/h
Dublin (IE) / eu-west-1
3.298533 USD/h 16.2266 USD/h
Mumbai (IN) / ap-south-1
2.66395 USD/h 16.8757 USD/h
Sydney (AU) / ap-southeast-2
16.8757 USD/h 16.8757 USD/h
Paris (FR) / eu-west-3
2.0805 USD/h 18.1738 USD/h
Singapore (SG) / ap-southeast-1
10.90995 USD/h 18.1738 USD/h
London (GB) / eu-west-2
1.9472 USD/h 19.4719 USD/h
Tokyo (JP) / ap-northeast-1
19.4719 USD/h 19.4719 USD/h
Frankfurt (DE) / eu-central-1
19.4719 USD/h 19.4719 USD/h
Sao Paulo (BR) / sa-east-1
2.2068 USD/h 22.0681 USD/h

Processor

vCPUs
192
Hypervisor
nitro
CPU Allocation
Dedicated
CPU Cores
96
CPU Speed
3.6 GHz
CPU Architecture
x86_64
CPU Manufacturer
AMD
CPU Family
EPYC
CPU Model
7R13
CPU L1D Cache
32 KiB
CPU L1D Cache Total
3 MiB
CPU L1I Cache
32 KiB
CPU L1I Cache Total
3 MiB
CPU L2 Cache
512 KiB
CPU L2 Cache Total
48 MiB
CPU L3 Cache
32 MiB
CPU L3 Cache Total
384 MiB
CPU Flags
fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, mmxext, fxsr_opt, pdpe1gb, rdtscp, lm, constant_tsc, rep_good, nopl, nonstop_tsc, cpuid, extd_apicid, aperfmperf, tsc_known_freq, pni, pclmulqdq, monitor, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, cmp_legacy, cr8_legacy, abm, sse4a, misalignsse, 3dnowprefetch, topoext, perfctr_core, ssbd, ibrs, ibpb, stibp, vmmcall, fsgsbase, bmi1, avx2, smep, bmi2, invpcid, rdseed, adx, smap, clflushopt, clwb, sha_ni, xsaveopt, xsavec, xgetbv1, clzero, xsaveerptr, rdpru, wbnoinvd, arat, npt, nrip_save, vaes, vpclmulqdq, rdpid
Ecpus
95.9
Scalability
99.9

System Resources and Accelerators

MEMORY
Memory Amount
768 GiB
Memory Amount Actual
768 GiB
Memory Generation
DDR4
Memory Speed
3200 Mhz
GPU
GPU Count
0
GPU Memory Min
0 MiB
GPU Memory Total
0 MiB
GPUs
    STORAGE
    Storage Size
    0 GB
    Storages
      NETWORK
      Network Speed Baseline
      100 Gbps
      Network Speed Max
      100 Gbps
      Network Storage Speed Baseline
      60 Gbps
      Network Storage Speed Max
      60 Gbps
      Inbound Traffic
      0 GB/month
      Outbound Traffic
      0 GB/month
      IPv4
      0

      Server Description

      A high-density compute instance featuring 192 vCPUs, 100 Gbps networking, and top-tier multi-core performance for demanding parallel workloads.

      General PurposeCompute Optimized

      Amazon Web Services inf2.48xlarge is an Inferentia Gen2 instance featuring 192 vCPUs and 96 physical cores powered by an AMD EPYC 7R13 processor running at 3.6 GHz. Built on the AWS Nitro hypervisor, it includes 768.0 GB of DDR4 memory and a 100 Gbps baseline network interface, with no local storage. The instance delivers top-tier multi-core CPU performance, ranking in the top 10% on stress-ng and PassMark CPU tests, alongside strong cached memory bandwidth and high-throughput Redis SET operations. However, its PassMark memory score is in the bottom 10%. The inf2.48xlarge is highly effective for multi-threaded compute tasks, software compilation, ray tracing, and high-bandwidth network services, though storage-dependent workloads must rely on network-attached volumes.

      Economics

      Average Price per Region

      Prices per Zone

      Lowest Prices

      Performance

      Workload Profiles

      4.45Score
      Precomputed compound score for Cache Intensive workloads. A weighted average (geometric mean) of benchmark scores compared to their medians: score = ∏ (x_i / m_i)^(w_i / Σw). The score of 1.0 represents a synthetic baseline server with the median performance of each component benchmark; 0.5 means roughly half the performance; and 2.0 means twice the performance of that reference profile. Component weights: 50% Redis RPS (pipeline=1, SET), 20% Redis RPS (pipeline=16, SET), 10% PassMark Memory Mark (composite), 10% Memory bandwidth (read, 16 MB ~ L3), 10% PassMark single-thread CPU. Rationale for component selection: In-memory key-value store workload, mixing direct Redis performance metrics with memory speed and latency benchmarks, and single-core CPU performance profiles.
      Component Score Weight Impact
      Raw Reference Normalized Target Actual
      Redis RPS (pipeline=1, SET) 15,350,843 ops/sec 1,896,415 ops/sec 8.09 50.00% 50.00% +184.00%
      Redis RPS (pipeline=16, SET) 106,276,782 ops/sec 11,539,523 ops/sec 9.21 20.00% 20.00% +55.90%
      PassMark Memory Mark (composite) 1,174 2,411 0.487 10.00% 10.00% -6.94%
      Memory bandwidth (read, 16 MB ~ L3) 273,867 MB/sec 107,368 MB/sec 2.55 10.00% 10.00% +9.81%
      PassMark single-thread CPU 1,980 Mops/s 2,363 Mops/s 0.838 10.00% 10.00% -1.75%

      Memory Bandwidth

      Compression

      OpenSSL

      Geekbench Single-Core

      Score: 1,677

      Geekbench Multi-Core

      Score: 15,331

      Passmark CPU Scores

      BENCHMARKSCORE
      Mark
      95525
      Compression
      2982357
      Encryption
      221586
      Extended Instructions
      153692
      Floating Point Maths
      483941
      Integer Maths
      880361
      Physics
      14535
      Prime Numbers
      1050
      Single Threaded
      1980
      String Sorting
      367059

      Passmark Memory Scores

      BENCHMARKSCORE
      Memory Mark
      1174
      Database Operations
      16815
      Memory Latency
      73
      Memory Read Cached
      10071
      Memory Read Uncached
      3053
      Memory Write
      2914

      Stress-ng Raw Scores

      Stress-ng Relative Multicore Performance

      LLM Inference Speed for Prompt Processing

      LLM Inference Speed for Text Generation

      Static Web Server

      Redis

      Alternatives

      Servers of the Same Family

      INSTANCEvCPUsMEMORYGPUs
      inf2.xlarge 416 GiB0
      inf2.8xlarge 32128 GiB0
      inf2.24xlarge 96384 GiB0

      Similar Servers

      INSTANCEVENDORvCPUsMEMORYGPUs

      inf2.48xlarge FAQs