vcg-a40-8c-40g-16vram by Vultr
(All-cores)
(Single-core)
Specifications
Server Metadata
Vendor ID | vultr |
Name | vcg-a40-8c-40g-16vram |
Description | Cloud GPU (8 vCPUs, 40.0 GiB RAM, 740 GB NVMe, 0.3333xA40 48 GiB VRAM) |
Family | Cloud GPU |
Hw Virt | |
Average Time To Start | 69 |
Status | active |
Observed At | 2026-07-20T13:40:44.473168 |
Availability
| REGION / ID | SPOT | ONDEMAND |
|---|
Processor
vCPUs | 8 |
CPU Allocation | Shared |
CPU Cores | 4 |
CPU Speed | 2 GHz |
CPU Architecture | x86_64 |
CPU Manufacturer | Intel |
CPU L1D Cache | 32 KiB |
CPU L1D Cache Total | 256 KiB |
CPU L1I Cache | 32 KiB |
CPU L1I Cache Total | 256 KiB |
CPU L2 Cache | 4 MiB |
CPU L2 Cache Total | 16 MiB |
CPU L3 Cache | 16 MiB |
CPU L3 Cache Total | 16 MiB |
CPU Flags | fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, rdtscp, lm, constant_tsc, rep_good, nopl, xtopology, cpuid, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, tsc_deadline_timer, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, abm, cpuid_fault, pti, ssbd, ibrs, ibpb, fsgsbase, bmi1, avx2, smep, bmi2, erms, invpcid, xsaveopt, arat |
Ecpus | 7.9 |
Scalability | 197.5 |
System Resources and Accelerators
| MEMORY | |
|---|---|
Memory Amount | 40 GiB |
Memory Amount Actual | 40 GiB |
| GPU | |
|---|---|
GPU Count | 0.3333 |
GPU Memory Min | 16 GiB |
GPU Memory Total | 16 GiB |
GPU Manufacturer | NVIDIA |
GPU Family | Ampere |
GPU Model | A40 |
GPUs |
| STORAGE | |
|---|---|
Storage Size | 740 GB |
Storage Type | nvme ssd |
Storages |
| NETWORK | |
|---|---|
Inbound Traffic | 0 GB/month |
Outbound Traffic | 8192 GB/month |
IPv4 | 1 |
CPU and System Topology
Server Description
A fractional GPU instance combining 16 GB VRAM, shared Intel processors, and NVMe storage for cost-efficient machine learning inference.
Vultr vcg-a40-8c-40g-16vram is a Cloud GPU server instance configured with a fractional NVIDIA Ampere A40 GPU delivering 16 GB of VRAM. It features 8 shared Intel x86_64 vCPUs running at 2.0 GHz, 40.0 GB of system memory, and a 740 GB NVMe SSD. Benchmark results show average performance for general CPU and memory operations, including a PassMark CPU score of 14090.9 and a multi-core stress-ng score of 14175.12. However, the instance achieves top-tier results in LLM inference, processing 3147.67 tokens/sec for 7B model prompts. The fractional GPU design provides a cost-efficient option for GPU-accelerated workloads, though the shared CPU allocation may introduce resource contention. This instance is suited for small-to-medium LLM inference, database operations, and GPU-accelerated web serving.
Economics
Performance
Memory Bandwidth
Compression
OpenSSL
Passmark CPU Scores
| BENCHMARK | SCORE |
|---|---|
Mark | 14091 |
Compression | 174552 |
Encryption | 4817 |
Extended Instructions | 11615 |
Floating Point Maths | 31322 |
Integer Maths | 39430 |
Physics | 2083 |
Prime Numbers | 126 |
Single Threaded | 2279 |
String Sorting | 23464 |
Passmark Memory Scores
| BENCHMARK | SCORE |
|---|---|
Memory Mark | 2190 |
Database Operations | 3659 |
Memory Latency | 65 |
Memory Read Cached | 22208 |
Memory Read Uncached | 9366 |
Memory Write | 10040 |
Stress-ng Raw Scores
Stress-ng Relative Multicore Performance
LLM Inference Speed for Prompt Processing
LLM Inference Speed for Text Generation
Static Web Server
Redis
Alternatives
Servers of the Same Family
| INSTANCE | vCPUs | MEMORY | GPUs |
|---|---|---|---|
| vcg-a40-1c-5g-2vram | 1 | 5 GiB | 1⁄24 |
| vcg-a40-2c-10g-4vram | 2 | 10 GiB | 1⁄12 |
| vcg-a16-2c-8g-2vram | 2 | 8 GiB | ⅛ |
| vcg-a16-2c-16g-4vram | 2 | 16 GiB | ¼ |
| vcg-a16-3c-32g-8vram | 3 | 32 GiB | ½ |
| vcg-a40-4c-20g-8vram | 4 | 20 GiB | ⅙ |
| vcg-a40-6c-30g-12vram | 6 | 30 GiB | ¼ |
Similar Servers
| INSTANCE | VENDOR | vCPUs | MEMORY | GPUs |
|---|