GPU Rental PricesGPU Rental Prices / GPU prices
Download data

A40 rental prices: cheapest $/hr today

The cheapest on-demand A40 rental today is $0.49/hr on RunPod, updated 2026-10-06. 48 GB VRAM, NVIDIA.

$0.49/hr
cheapest on-demand rate · RunPod (secure cloud) · source: www.runpod.io
Rent on RunPod →
secure cloud · runpod · $0.49/hr · 2026-10-06on-demand · vultr · $1.71/hr · 2026-10-06
$0.49/hr$1.71/hr
2 offers · dashed line: $1.10 median per offer · 2026-10-06. ● On-demand · ◆ Spot/community · ■ Reserved · ○ Serverless.

Pricing at a glance

$0.49
cheapest on-demand $/hr
$1.10
median on-demand $/hr, each provider counted once
$1.71
max on-demand $/hr
2
live offers
2
providers

All current A40 offers

Provider, memory & availability

2 offers. Rates are per GPU; total is for the smallest listed node.

GPU / providerBilling$/GPU-hourGPUs / total per hourAvailability / details
A40RunPod48 GB VRAMsecure cloud$0.49per GPU-hour+0.0% over 30 daysNode minimum and total not reportedNot reportedView provider
2026-10-06Details
CPU
9 vCPU
RAM
50 GB RAM
Disk
Disk not reported
Region
Not reported
Updated
2026-10-06

Source price ↗ · Report this price

A40Vultr48 GB VRAMon-demand$1.71per GPU-hour+0.0% over 30 days1×
$1.71per node-hour
Not reportedView provider
2026-10-06Details
CPU
vCPU not reported
RAM
RAM not reported
Disk
Disk not reported
Region
Not reported
Updated
2026-10-06

Source price ↗ · Report this price

Tables rank by price, lowest first. Commercial relationships never affect ordering. How we collect and rank prices.

On-demand and secure rates do not require a reservation. Community and spot capacity can be interrupted. Prices are per GPU per hour, before storage/egress. Every row shows its update date.

A40 hardware specs

A40 hardware specifications
ArchitectureAmpere
VRAM48 GB
FP16 TFLOPS (dense)150
Memory bandwidth696 GB/s
TDP300 W
InterconnectNVLink 112.5GB/s
Launched2020

Datasheet spec for the A40 variant. Source: images.nvidia.com.

Compare A40 configurations

A40 purchase prices · Rent or buy calculator · Compare two GPUs

LLM memory estimates for this GPU

Qwen2.5 14B at 16-bit weights
36.4GB
Weights 29.5 GB · cache 0.8 GB · headroom 6.1 GB. Capacity line: 48 GB.

These models have an estimated inference footprint within 48 GB at one or more precisions. The scenario uses 4,096 tokens, batch 1, a 16-bit KV cache and 20% runtime headroom. A memory fit does not establish speed or software compatibility.

Model16-bit GB8-bit GB4-bit GB
Qwen2.5 0.5B1.2 ✓0.7 ✓0.4 ✓
Qwen2.5 1.5B3.8 ✓2.0 ✓1.1 ✓
Qwen2.5 3B7.6 ✓3.9 ✓2.0 ✓
Qwen2.5 7B18.6 ✓9.4 ✓4.9 ✓
Qwen2.5 14B36.4 ✓18.7 ✓9.8 ✓
Qwen2.5 32B79.9 40.6 ✓20.9 ✓
Qwen2.5 72B176.1 88.9 45.2 ✓
Qwen2.5 Coder 0.5B1.2 ✓0.7 ✓0.4 ✓
Qwen2.5 Coder 1.5B3.8 ✓2.0 ✓1.1 ✓
Qwen2.5 Coder 3B7.6 ✓3.9 ✓2.0 ✓
Qwen2.5 Coder 7B18.6 ✓9.4 ✓4.9 ✓
Qwen2.5 Coder 14B36.4 ✓18.7 ✓9.8 ✓
Qwen2.5 Coder 32B79.9 40.6 ✓20.9 ✓
Qwen3 0.6B2.4 ✓1.5 ✓1.0 ✓
Qwen3 1.7B5.4 ✓3.0 ✓1.8 ✓
Qwen3 4B10.4 ✓5.6 ✓3.1 ✓
Qwen3 8B20.4 ✓10.6 ✓5.6 ✓
Qwen3 14B36.2 ✓18.5 ✓9.7 ✓
Qwen3 32B79.9 40.6 ✓20.9 ✓
Qwen3 30B A3B73.8 37.1 ✓18.8 ✓
DeepSeek R1 Distill Qwen 1.5B4.4 ✓2.3 ✓1.2 ✓
DeepSeek R1 Distill Qwen 7B18.6 ✓9.4 ✓4.9 ✓
DeepSeek R1 Distill Qwen 14B36.4 ✓18.7 ✓9.8 ✓
DeepSeek R1 Distill Qwen 32B79.9 40.6 ✓20.9 ✓
DeepSeek R1 Distill Llama 8B19.9 ✓10.3 ✓5.5 ✓
DeepSeek R1 Distill Llama 70B170.9 86.3 43.9 ✓
Mistral 7B v0.318.0 ✓9.3 ✓5.0 ✓
Mixtral 8x7B v0.1112.7 56.7 28.7 ✓
Mistral Small 24B 250157.4 29.1 ✓14.9 ✓
Mistral Nemo 240730.2 ✓15.5 ✓8.2 ✓
phi 436.2 ✓18.6 ✓9.8 ✓

Compare model memory requirements or change context and memory assumptions on a model's page.

When to choose A40

Use the model memory estimates above to check whether your weights, context and batch size fit within 48 GB. Then compare the exact variant, node size and software support. Capacity alone does not establish inference speed or training throughput.

If your workload fits within 32 GB, compare RTX PRO 4500 rental prices from $0.72/GPU-hour. The smaller memory budget is the trade-off; price alone does not make the cards interchangeable.

A40 price changes

Changes use the geometric mean of price ratios for the same provider, variant, tier, GPU count, instance and region. New listings are not counted as price changes.

Price history

on-demand rate per day · median on-demand rate per day (dashed), since 2026-07-05: $0.25 low · $0.49 high. Full ledger on the history hub.

FAQ

How much does it cost to rent an A40?

The cheapest A40 rental we have is $0.49 per GPU-hour on RunPod (secure cloud), updated 2026-10-06.

How much VRAM does an A40 have?

48 GB.

Where is the A40 cheapest right now?

RunPod has the cheapest verified on-demand rate today at $0.49 per GPU-hour. Verified 2026-10-06.

Related

All 48 GB+ GPUs · GPUs under $1/hr · Price movers this week · RunPod review · L40S vs A40 · A100 vs A40 · A10 vs A40

Should you rent or buy?

Compare A40 purchase prices and quote availability, then enter your usage and electricity costs in the rent-versus-buy calculator. Purchase listings have their own variants, stock status and update dates.

Compare your offers

Saved prices keep their original dates. Confirm the configuration, terms and availability at the source.

Select up to three offers to compare their rates and terms.