vcg-a16-2c-16g-4vram by Vultr
(All-cores)
(Single-core)
Specifications
Server Metadata
Vendor ID | vultr |
Name | vcg-a16-2c-16g-4vram |
Description | Cloud GPU (2 vCPUs, 16.0 GiB RAM, 80 GB NVMe, 0.25xA16 16 GiB VRAM) |
Family | Cloud GPU |
Hw Virt | |
Average Time To Start | 54 |
Status | active |
Observed At | 2026-07-20T13:40:44.472936 |
Availability
| REGION / ID | SPOT | ONDEMAND |
|---|
Processor
vCPUs | 2 |
CPU Allocation | Shared |
CPU Cores | 1 |
CPU Speed | 2 GHz |
CPU Architecture | x86_64 |
CPU Manufacturer | Intel |
CPU L1D Cache | 32 KiB |
CPU L1D Cache Total | 64 KiB |
CPU L1I Cache | 32 KiB |
CPU L1I Cache Total | 64 KiB |
CPU L2 Cache | 4 MiB |
CPU L2 Cache Total | 4 MiB |
CPU L3 Cache | 16 MiB |
CPU L3 Cache Total | 16 MiB |
CPU Flags | fpu, vme, de, pse, tsc, msr, pae, mce, cx8, apic, sep, mtrr, pge, mca, cmov, pat, pse36, clflush, mmx, fxsr, sse, sse2, ht, syscall, nx, rdtscp, lm, constant_tsc, rep_good, nopl, xtopology, cpuid, tsc_known_freq, pni, pclmulqdq, ssse3, fma, cx16, pcid, sse4_1, sse4_2, x2apic, movbe, popcnt, tsc_deadline_timer, aes, xsave, avx, f16c, rdrand, hypervisor, lahf_lm, abm, cpuid_fault, pti, ssbd, ibrs, ibpb, fsgsbase, bmi1, avx2, smep, bmi2, erms, invpcid, xsaveopt, arat |
Ecpus | 2 |
Scalability | 200 |
System Resources and Accelerators
| MEMORY | |
|---|---|
Memory Amount | 16 GiB |
Memory Amount Actual | 16 GiB |
| GPU | |
|---|---|
GPU Count | 0.25 |
GPU Memory Min | 4 GiB |
GPU Memory Total | 4 GiB |
GPU Manufacturer | NVIDIA |
GPU Family | Ampere |
GPU Model | A16 |
GPUs |
| STORAGE | |
|---|---|
Storage Size | 80 GB |
Storage Type | nvme ssd |
Storages |
| NETWORK | |
|---|---|
Inbound Traffic | 0 GB/month |
Outbound Traffic | 2048 GB/month |
IPv4 | 1 |
CPU and System Topology
Server Description
A cost-efficient fractional GPU instance combining shared Intel processors with NVIDIA Ampere acceleration for entry-level machine learning and development workloads.
Vultr vcg-a16-2c-16g-4vram is a fractional GPU instance designed for entry-level accelerated workloads. It is equipped with 2 shared Intel vCPUs, 16.0 GB of system memory, and an 80 GB NVMe SSD. Acceleration is provided by a 0.25 slice of an NVIDIA Ampere A16 GPU with 4 GB of VRAM. While CPU-bound benchmarks and memory bandwidth are relatively weak—as shown by a passmark CPU score of 3681.75—the instance delivers top-tier performance in LLM prompt processing, reaching 308.92 tokens per second on a 7B model. This profile offers a cost-efficient entry point for GPU development, but is poorly suited for database operations or heavy multi-threaded CPU tasks. It is best utilized for lightweight machine learning inference, small model testing, and GPU-accelerated application development.
Economics
Performance
Memory Bandwidth
Compression
OpenSSL
Passmark CPU Scores
| BENCHMARK | SCORE |
|---|---|
Mark | 3682 |
Compression | 40790 |
Encryption | 1191 |
Extended Instructions | 2751 |
Floating Point Maths | 8014 |
Integer Maths | 9757 |
Physics | 531 |
Prime Numbers | 31 |
Single Threaded | 2239 |
String Sorting | 5772 |
Passmark Memory Scores
| BENCHMARK | SCORE |
|---|---|
Memory Mark | 2014 |
Database Operations | 1006 |
Memory Latency | 51 |
Memory Read Cached | 22739 |
Memory Read Uncached | 11827 |
Memory Write | 10850 |
Stress-ng Raw Scores
Stress-ng Relative Multicore Performance
LLM Inference Speed for Prompt Processing
LLM Inference Speed for Text Generation
Static Web Server
Redis
Alternatives
Servers of the Same Family
| INSTANCE | vCPUs | MEMORY | GPUs |
|---|---|---|---|
| vcg-a40-1c-5g-2vram | 1 | 5 GiB | 1⁄24 |
| vcg-a40-2c-10g-4vram | 2 | 10 GiB | 1⁄12 |
| vcg-a16-2c-8g-2vram | 2 | 8 GiB | ⅛ |
| vcg-a16-3c-32g-8vram | 3 | 32 GiB | ½ |
| vcg-a40-4c-20g-8vram | 4 | 20 GiB | ⅙ |
| vcg-a40-6c-30g-12vram | 6 | 30 GiB | ¼ |
| vcg-a16-6c-64g-16vram | 6 | 64 GiB | 1 |
Similar Servers
| INSTANCE | VENDOR | vCPUs | MEMORY | GPUs |
|---|