vcg-a40-4c-20g-8vram by Vultr
(All-cores)
(Single-core)
Specifications
Server Metadata
Vendor ID | vultr |
Name | vcg-a40-4c-20g-8vram |
Description | Cloud GPU (4 vCPUs, 20.0 GiB RAM, 360 GB NVMe, 0.1667xA40 48 GiB VRAM) |
Family | Cloud GPU |
Hw Virt | |
Average Time To Start | 77.67 |
Status | active |
Observed At | 2026-07-20T16:35:16.793899 |
Availability
| REGION / ID | SPOT | ONDEMAND |
|---|---|---|
Frankfurt (DE) / fra | - | 0.288 USD/h |
Processor
vCPUs | 4 |
CPU Allocation | Shared |
CPU Cores | 2 |
CPU Speed | 2 GHz |
CPU Architecture | x86_64 |
CPU Manufacturer | Intel |
CPU L1D Cache | 32 KiB |
CPU L1D Cache Total | 128 KiB |
CPU L1I Cache | 32 KiB |
CPU L1I Cache Total | 128 KiB |
CPU L2 Cache | 4 MiB |
CPU L2 Cache Total | 8 MiB |
CPU L3 Cache | 16 MiB |
CPU L3 Cache Total | 16 MiB |
CPU Flags | fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, rdtscp, lm, constant_tsc, rep_good, nopl, xtopology, cpuid, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, tsc_deadline_timer, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, abm, cpuid_fault, pti, ssbd, ibrs, ibpb, fsgsbase, bmi1, avx2, smep, bmi2, erms, invpcid, xsaveopt, arat |
Ecpus | 4 |
Scalability | 200 |
System Resources and Accelerators
| MEMORY | |
|---|---|
Memory Amount | 20 GiB |
Memory Amount Actual | 20 GiB |
| GPU | |
|---|---|
GPU Count | 0.1667 |
GPU Memory Min | 8 GiB |
GPU Memory Total | 8 GiB |
GPU Manufacturer | NVIDIA |
GPU Family | Ampere |
GPU Model | A40 |
GPUs |
| STORAGE | |
|---|---|
Storage Size | 360 GB |
Storage Type | nvme ssd |
Storages |
| NETWORK | |
|---|---|
Inbound Traffic | 0 GB/month |
Outbound Traffic | 5120 GB/month |
IPv4 | 1 |
CPU and System Topology
Server Description
A fractional GPU instance combining Intel processors and NVIDIA Ampere acceleration for efficient, small-scale machine learning inference and development.
Vultr vcg-a40-4c-20g-8vram is a fractional GPU instance designed for specialized workloads requiring hardware acceleration without the cost of a full GPU. It features 4 shared Intel x86_64 vCPUs running at 2.0 GHz, 20.0 GB of RAM, and a 360 GB NVMe SSD. Graphics acceleration is provided by a fractional 0.1667 NVIDIA Ampere A40 GPU with 8 GB of VRAM. Benchmark data highlights top-tier performance in LLM inference for both small and medium models, while general CPU and database operations perform in the average to weak ranges. This resource profile makes the instance highly cost-efficient for lightweight machine learning inference, model prototyping, and GPU-accelerated development, though it is less suited for heavy database operations or intensive multi-threaded CPU tasks.
Economics
Performance
Memory Bandwidth
Compression
OpenSSL
Passmark CPU Scores
| BENCHMARK | SCORE |
|---|---|
Mark | 7273 |
Compression | 83753 |
Encryption | 2400 |
Extended Instructions | 5705 |
Floating Point Maths | 15742 |
Integer Maths | 19643 |
Physics | 1066 |
Prime Numbers | 63 |
Single Threaded | 2236 |
String Sorting | 11564 |
Passmark Memory Scores
| BENCHMARK | SCORE |
|---|---|
Memory Mark | 2219 |
Database Operations | 1916 |
Memory Latency | 51 |
Memory Read Cached | 23004 |
Memory Read Uncached | 10806 |
Memory Write | 8898 |
Stress-ng Raw Scores
Stress-ng Relative Multicore Performance
LLM Inference Speed for Prompt Processing
LLM Inference Speed for Text Generation
Static Web Server
Redis
Alternatives
Servers of the Same Family
| INSTANCE | vCPUs | MEMORY | GPUs |
|---|---|---|---|
| vcg-a40-1c-5g-2vram | 1 | 5 GiB | 1⁄24 |
| vcg-a40-2c-10g-4vram | 2 | 10 GiB | 1⁄12 |
| vcg-a16-2c-8g-2vram | 2 | 8 GiB | ⅛ |
| vcg-a16-2c-16g-4vram | 2 | 16 GiB | ¼ |
| vcg-a16-3c-32g-8vram | 3 | 32 GiB | ½ |
| vcg-a40-6c-30g-12vram | 6 | 30 GiB | ¼ |
| vcg-a16-6c-64g-16vram | 6 | 64 GiB | 1 |
Similar Servers
| INSTANCE | VENDOR | vCPUs | MEMORY | GPUs |
|---|