$GPU Rental Prices.com
verified 2026-09-02

Rent a GPU server: dedicated, VPS or GPU-as-a-service

To rent GPU compute, first choose the product shape: a full dedicated node for tightly connected multi-GPU work, a single-GPU cloud instance for flexible server access, serverless for bursty jobs, or interruptible marketplace capacity for checkpointable work. The comparison below maps those choices to live offers from34 providers, checked daily and linked to their sources.

The vocabulary, mapped to real products

Dedicated / bare-metal GPU server

A whole physical machine, no virtualization, typically an 8-GPU node with NVLink or InfiniBand for training. In our index these are the prices badged with a multi-GPU minimum: AWS, CoreWeave, Latitude.sh, Microsoft Azure, OVHcloud, Scaleway, SF Compute, Vultr currently price capacity this way (per-GPU rate, full node required).

GPU VPS / cloud server with GPU

A virtualized instance with a dedicated GPU passed through: rentable one GPU at a time, billed hourly or finer. This is the bulk of the market; 25 of our tracked providers sell single-GPU on-demand instances. The table below shows the cheapest such rate per GPU model.

GPU-as-a-service (GPUaaS)

The umbrella term for renting GPU compute instead of buying hardware; at its most abstract, serverless platforms (Baseten, Cerebrium, fal, Fireworks AI, Koyeb, Modal, Replicate) bill per second of compute with no VM at all. Full comparison on the serverless pricing page.

Marketplace / interruptible capacity

Spot and community-cloud capacity: cheaper than any hosting contract, but reclaimable with little warning. Currently listed by AWS, Google Cloud, Microsoft Azure, Prime Intellect, RunPod, SaladCloud, SF Compute, Spheron, TensorDock, Vultr. Good for checkpointable jobs; wrong for a production endpoint.

Cheapest GPU server rental per model, verified 2026-09-02

GPUVRAMOn-demand $/hr~$/mo (730h)Cheapest atVerified
RTX A400016 GB$0.10$73TensorDock on-demand2026-09-02Rent →
Tesla V10016 GB$0.17$124DataCrunch on-demand2026-09-02Rent →
RTX 309024 GB$0.20$146TensorDock on-demand2026-09-02Rent →
RTX A500024 GB$0.27$197RunPod secure cloud2026-09-02Rent →
RTX 409024 GB$0.35$255TensorDock on-demand2026-09-02Rent →
Tesla T416 GB$0.35$255Google Cloud on-demand2026-09-02Rent →
RTX A600048 GB$0.35$255Thunder Compute on-demand2026-09-02Rent →
A3024 GB$0.35$255Massed Compute on-demand2026-09-02Rent →
L424 GB$0.44$321Jarvislabs on-demand2026-09-02Rent →
A4048 GB$0.44$321RunPod secure cloud2026-09-02Rent →
Quadro RTX 500016 GB$0.60$438OVHcloud on-demand2026-09-02Rent →
Quadro RTX 600024 GB$0.69$504Lambda on-demand2026-09-02Rent →
RTX PRO 450032 GB$0.72$526RunPod secure cloud2026-09-02Rent →
RTX 6000 Ada48 GB$0.75$548TensorDock on-demand2026-09-02Rent →
L4048 GB$0.79$577Thunder Compute on-demand2026-09-02Rent →
RTX 509032 GB$0.86$628Spheron on-demand2026-09-02Rent →
L40S48 GB$0.88$642Massed Compute on-demand2026-09-02Rent →
A10040-80 GB$0.89$650Jarvislabs on-demand2026-09-02Rent →
A1024 GB$1.00$730OVHcloud on-demand2026-09-02Rent →
RTX PRO 600096 GB$1.10$800Google Cloud on-demand2026-09-02Rent →
H10080-94 GB$1.99$1,453Voltage Park on-demand2026-09-02Rent →
GH20096 GB$2.29$1,672Lambda on-demand2026-09-02Rent →
MI300X192 GB$2.39$1,745RunPod secure cloud2026-09-02Rent →
H200141-143 GB$2.60$1,898GMI Cloud on-demand2026-09-02Rent →
B200180 GB$3.49$2,548Prime Intellect on-demand2026-09-02Rent →
MI325X256 GB$3.80$2,774DigitalOcean on-demand2026-09-02Rent →
B300288 GB$4.99$3,643Prime Intellect on-demand2026-09-02Rent →
MI350X288 GB$6.16$4,497DigitalOcean on-demand2026-09-02Rent →
GB200186 GB$8.00$5,840GMI Cloud on-demand2026-09-02Rent →
GB300288 GB$8.62$6,293DataCrunch on-demand2026-09-02Rent →
RTX 508016 GBspot onlySaladCloud community cloud2026-09-02Rent →
MI355X288 GBspot onlyVultr spot / interruptible 8× min2026-09-02Rent →

Tables rank by price, lowest first. Commercial relationships never affect ordering.How we collect and rank prices.

Monthly figures are computed, not quoted: on-demand $/hr × 730 hours (an always-on month), before storage and egress. Per-second billing means real bills are lower at partial utilization; run your own numbers in the rent-vs-buy calculator. Prices badged with a multi-GPU minimum require renting the whole node. Click any GPU for its full provider table and price history, or browse GPUs under $1/hr and 80 GB+ VRAM options.

Picking the right shape of GPU server

Rule of thumb from the data: rent a dedicated full node when you need multi-GPU training with fast interconnect (the per-GPU rate is often lower at the node level, but you pay for 8 GPUs whether you saturate them or not). Rent an on-demand instance for everything single-GPU: fine-tuning, inference serving, rendering, experiments. Use serverless when traffic is bursty and idle hours would dominate the bill. Use spot or community capacity when the job checkpoints and a restart costs you minutes, not customers. The spot vs on-demand vs serverless guide walks through the decision, and the fees matrix covers what the hourly sticker hides.

FAQ

What is a GPU server?

A GPU server is a machine with one or more GPUs attached, rented for compute-heavy work: AI training and inference, rendering, simulation. In practice the term covers three different products. A dedicated or bare-metal GPU server gives you a whole physical machine. A GPU VPS or cloud GPU server gives you a virtualized instance with a GPU attached, rentable by the hour. GPU-as-a-service abstracts the server away entirely and bills per second of compute. All three run the same chips; they differ in isolation, billing granularity and price.

What is the difference between a dedicated GPU server and a cloud GPU server?

A dedicated (bare-metal) GPU server is a physical machine reserved for you alone: full hardware access, no virtualization overhead, and usually a full node of 8 GPUs with fast interconnect for training. In our data these are the offers with a multi-GPU minimum, currently sold by AWS, CoreWeave, Latitude.sh, Microsoft Azure, OVHcloud, Scaleway, SF Compute, Vultr. A cloud GPU server (or GPU VPS) is a virtualized instance carved from shared hardware, rentable one GPU at a time by the hour, which is what most of the 25 single-GPU providers in our index sell. Dedicated wins on isolation and interconnect; cloud instances win on flexibility and entry price.

How much does a GPU server cost per month?

At today's verified rates (2026-09-02), running a server around the clock costs its hourly rate times 730 hours. The cheapest on-demand GPU in our index, the RTX A4000 at $0.10/hr, works out to about $73/month; an H100 at the cheapest on-demand rate of $1.99/hr works out to about $1,453/month. That assumes 100% uptime: per-second and per-minute billing means you only pay for hours you actually use, and spot or community capacity cuts the rate further if your work tolerates interruption. Storage and egress are billed separately by many providers.

What is a GPU VPS?

A GPU VPS (virtual private server with a GPU) is a virtualized slice of a physical machine with a GPU passed through to it: you get root access and a dedicated GPU, but share the underlying host with other tenants. Hosting companies say GPU VPS; AI clouds say on-demand GPU instance or cloud GPU. They are the same product category, and the on-demand prices on this page are exactly that: virtualized instances with dedicated GPUs, billed by the hour or finer.

Related

All live GPU prices · What neoclouds are and what they charge · All cloud GPUs by price · Serverless GPU pricing · Rent vs buy calculator · Cheapest cloud GPU, with full-cost checks · Hidden fees and egress matrix · GPU rental glossary · All providers