GPU Rental PricesGPU Rental Prices / GPU prices
Download data
updated 2026-10-06

Runpod serverless pricing and other GPU clouds

Compare published per-second GPU rates with hourly instance prices. Idle workers, cold starts and minimum billing rules can change the bill; the label “serverless” does not establish those rules. Rates below show both units; the published unit and rounding differ by provider.

GPU$/hr equivalent$/secondProviderUpdated
Tesla T4$0.590$0.000164Modal2026-10-06Try →
Tesla T4$0.590$0.000164Cerebrium2026-10-06Try →
Tesla T4$0.631$0.000175Baseten2026-10-06Try →
L4$0.70$0.000194Koyeb2026-10-06Try →
RTX A6000$0.75$0.000208Koyeb2026-10-06Try →
L4$0.799$0.000222Modal2026-10-06Try →
L4$0.799$0.000222Cerebrium2026-10-06Try →
Tesla T4$0.81$0.000225Replicate2026-10-06Try →
L4$0.848$0.000236Baseten2026-10-06Try →
RTX 4090$1.10$0.000306RunPod2026-10-06Try →
A10$1.10$0.000306Cerebrium2026-10-06Try →
A10$1.10$0.000306Modal2026-10-06Try →
RTX PRO 4500 Blackwell$1.15$0.000319RunPod2026-10-06Try →
L40S$1.20$0.000333Koyeb2026-10-06Try →
A10$1.21$0.000335Baseten2026-10-06Try →
RTX 5090$1.58$0.000439RunPod2026-10-06Try →
A100 (unspecified)$1.60$0.000444Koyeb2026-10-06Try →
L40S$1.95$0.000542Modal2026-10-06Try →
L40S$1.95$0.000542Cerebrium2026-10-06Try →
A100 PCIe 40GB$2.00$0.000555Cerebrium2026-10-06Try →
A100 PCIe 40GB$2.10$0.000583Modal2026-10-06Try →
A100 (unspecified)$2.10$0.000583Cerebrium2026-10-06Try →
A100 SXM 80GB$2.15$0.000597Koyeb2026-10-06Try →
RTX PRO 6000 (unspecified)$2.20$0.000611Koyeb2026-10-06Try →
A100 (unspecified)$2.50$0.000694Modal2026-10-06Try →
RTX PRO 6000 (unspecified)$2.50$0.000694Cerebrium2026-10-06Try →
H100 (unspecified)$2.50$0.000694Koyeb2026-10-06Try →
A100 (unspecified)$2.72$0.000756RunPod2026-10-06Try →
H200 (SXM)$3.00$0.000833Koyeb2026-10-06Try →
RTX PRO 6000 (unspecified)$3.03$0.000842Modal2026-10-06Try →
H100 (unspecified)$3.40$0.000944Cerebrium2026-10-06Try →
RTX PRO 6000 (unspecified)$3.49$0.000969RunPod2026-10-06Try →
L40S$3.51$0.000975Replicate2026-10-06Try →
H100 SXM$3.95$0.001097Modal2026-10-06Try →
RTX PRO 6000 (unspecified)$4.00$0.001111fal2026-10-06Try →
A100 (unspecified)$4.00$0.001111Baseten2026-10-06Try →
H200 (SXM)$4.20$0.001166Cerebrium2026-10-06Try →
H100 (unspecified)$4.50$0.001250fal2026-10-06Try →
H200 (SXM)$4.54$0.001261Modal2026-10-06Try →
H100 (unspecified)$4.79$0.001331RunPod2026-10-06Try →
A100 (unspecified)$5.04$0.001400Replicate2026-10-06Try →
H100 (unspecified)$5.49$0.001525Replicate2026-10-06Try →
H200 (SXM)$5.49$0.001525Replicate2026-10-06Try →
B200$5.50$0.001528Koyeb2026-10-06Try →
H200 (SXM)$5.93$0.001647RunPod2026-10-06Try →
H200 (SXM)$6.00$0.001667fal2026-10-06Try →
B200$6.01$0.001670Cerebrium2026-10-06Try →
B200$6.25$0.001736Modal2026-10-06Try →
H100 (unspecified)$6.50$0.001806Baseten2026-10-06Try →
B300$7.10$0.001972Modal2026-10-06Try →
B200$7.99$0.002219fal2026-10-06Try →
H100 (unspecified)$8.00$0.002222Fireworks AI2026-10-06Try →
H200 (SXM)$8.00$0.002222Fireworks AI2026-10-06Try →
B200$8.64$0.002400RunPod2026-10-06Try →
B200$9.98$0.002772Baseten2026-10-06Try →
GB200$9.99$0.002775fal2026-10-06Try →
B300$12.99$0.003608fal2026-10-06Try →
B200$13.00$0.003611Fireworks AI2026-10-06Try →
B300$15.00$0.004167Fireworks AI2026-10-06Try →
GB300$20.00$0.005556Fireworks AI2026-10-06Try →

Runpod Flex workers are billed from startup until fully stopped, rounded up to the nearest second. That includes loading the model, request execution and the idle timeout. Active workers stay running and require a sales inquiry for discounted rates. The table includes only individually named GPUs; mixed-card classes cannot promise a particular card.

Runpod publishes hourly equivalents on its pricing page; their per-second values here are those figures divided by 3,600 and may be rounded. Modal's per-second GPU rates are multiplied by 3,600 for comparison. Separately billed CPU, RAM and storage are excluded.

When serverless wins

Pick serverless when

  • Traffic is bursty or unpredictable
  • You serve a model occasionally (demos, side projects, internal tools)
  • The provider’s scale-to-zero and billing rules match your idle periods

Pick an instance when

  • Utilization is sustained (training, batch jobs, steady APIs)
  • Cold starts are unacceptable
  • You need full control of the environment or multi-GPU nodes

Modal includes a recurring monthly free credit tier; see free credits. RunPod also offers serverless endpoints priced separately from the pod rates shown on the RunPod page.

When does an always-on GPU cost less?

For the same GPU variant, divide the single-GPU instance rate by the serverless hourly equivalent. That fraction of an hour is the compute-only crossover. Request overhead and separately billed CPU, RAM, storage or warm workers can move it.

GPUServerless $/secondAlways-on $/hourCrossover per hour
B200RunPod · $0.002400Lium · $5.4037.5 active minutes (62.5% duty)
H200 (SXM)RunPod · $0.001647DigitalOcean · $4.4745.2 active minutes (75.4% duty)
RTX PRO 6000 (unspecified)RunPod · $0.000969DataCrunch · $2.0835.8 active minutes (59.7% duty)
H100 (unspecified)RunPod · $0.001331Lium · $1.7521.9 active minutes (36.5% duty)
A100 (unspecified)RunPod · $0.000756Massed Compute · $1.3529.8 active minutes (49.6% duty)
RTX PRO 4500 BlackwellRunPod · $0.000319Massed Compute · $0.9248.0 active minutes (80.0% duty)
B300fal · $0.003608Lium · $8.0037.0 active minutes (61.6% duty)
B200fal · $0.002219Lium · $5.4040.6 active minutes (67.6% duty)
H200 (SXM)fal · $0.001667DigitalOcean · $4.4744.7 active minutes (74.5% duty)
H100 (unspecified)fal · $0.001250Lium · $1.7523.3 active minutes (38.9% duty)
RTX PRO 6000 (unspecified)fal · $0.001111DataCrunch · $2.0831.3 active minutes (52.1% duty)
B300Modal · $0.001972Lium · $8.00No crossover within one hour; this serverless rate is lower even at full utilization
B200Modal · $0.001736Lium · $5.4051.8 active minutes (86.4% duty)
H200 (SXM)Modal · $0.001261DigitalOcean · $4.4759.1 active minutes (98.5% duty)
H100 SXMModal · $0.001097Massed Compute · $3.1447.7 active minutes (79.5% duty)
RTX PRO 6000 (unspecified)Modal · $0.000842DataCrunch · $2.0841.3 active minutes (68.8% duty)
A100 (unspecified)Modal · $0.000694Massed Compute · $1.3532.4 active minutes (54.0% duty)
A100 PCIe 40GBModal · $0.000583Lambda · $1.9956.9 active minutes (94.8% duty)
L40SModal · $0.000542Massed Compute · $0.9729.8 active minutes (49.7% duty)
A10Modal · $0.000306OVHcloud · $1.0054.5 active minutes (90.8% duty)
L4Modal · $0.000222AWS · $0.805No crossover within one hour; this serverless rate is lower even at full utilization
A100 (unspecified)Replicate · $0.001400Massed Compute · $1.3516.1 active minutes (26.8% duty)
H100 (unspecified)Replicate · $0.001525Lium · $1.7519.1 active minutes (31.9% duty)
L40SReplicate · $0.000975Massed Compute · $0.9716.6 active minutes (27.6% duty)
H200 (SXM)Replicate · $0.001525DigitalOcean · $4.4748.9 active minutes (81.4% duty)
H100 (unspecified)Fireworks AI · $0.002222Lium · $1.7513.1 active minutes (21.9% duty)
H200 (SXM)Fireworks AI · $0.002222DigitalOcean · $4.4733.5 active minutes (55.9% duty)
B200Fireworks AI · $0.003611Lium · $5.4024.9 active minutes (41.5% duty)
B300Fireworks AI · $0.004167Lium · $8.0032.0 active minutes (53.3% duty)
GB300Fireworks AI · $0.005556DataCrunch · $10.3231.0 active minutes (51.6% duty)
L4Baseten · $0.000236AWS · $0.80556.9 active minutes (94.9% duty)
A10Baseten · $0.000335OVHcloud · $1.0049.7 active minutes (82.8% duty)
A100 (unspecified)Baseten · $0.001111Massed Compute · $1.3520.3 active minutes (33.8% duty)
H100 (unspecified)Baseten · $0.001806Lium · $1.7516.2 active minutes (26.9% duty)
B200Baseten · $0.002772Lium · $5.4032.5 active minutes (54.1% duty)
L4Cerebrium · $0.000222AWS · $0.805No crossover within one hour; this serverless rate is lower even at full utilization
A10Cerebrium · $0.000306OVHcloud · $1.0054.5 active minutes (90.8% duty)
A100 PCIe 40GBCerebrium · $0.000555Lambda · $1.9959.8 active minutes (99.6% duty)
L40SCerebrium · $0.000542Massed Compute · $0.9729.8 active minutes (49.7% duty)
A100 (unspecified)Cerebrium · $0.000583Massed Compute · $1.3538.6 active minutes (64.3% duty)
H100 (unspecified)Cerebrium · $0.000944Lium · $1.7530.9 active minutes (51.5% duty)
H200 (SXM)Cerebrium · $0.001166DigitalOcean · $4.47No crossover within one hour; this serverless rate is lower even at full utilization
B200Cerebrium · $0.001670Lium · $5.4053.9 active minutes (89.8% duty)
RTX PRO 6000 (unspecified)Cerebrium · $0.000694DataCrunch · $2.0850.1 active minutes (83.5% duty)
L4Koyeb · $0.000194AWS · $0.805No crossover within one hour; this serverless rate is lower even at full utilization
RTX A6000Koyeb · $0.000208Massed Compute · $0.5745.6 active minutes (76.0% duty)
L40SKoyeb · $0.000333Massed Compute · $0.9748.5 active minutes (80.8% duty)
A100 (unspecified)Koyeb · $0.000444Massed Compute · $1.3550.6 active minutes (84.4% duty)
A100 SXM 80GBKoyeb · $0.000597Massed Compute · $1.3838.5 active minutes (64.2% duty)
RTX PRO 6000 (unspecified)Koyeb · $0.000611DataCrunch · $2.0856.9 active minutes (94.8% duty)
H100 (unspecified)Koyeb · $0.000694Lium · $1.7542.0 active minutes (70.0% duty)
H200 (SXM)Koyeb · $0.000833DigitalOcean · $4.47No crossover within one hour; this serverless rate is lower even at full utilization
B200Koyeb · $0.001528Lium · $5.4058.9 active minutes (98.2% duty)

At a request duration you supply, requests at crossover = active seconds ÷ billed seconds per request. This assumes no overlap between requests and no idle worker charges. Per-second GPU pricing does not by itself guarantee zero idle charges.

Compare your offers

Saved prices keep their original dates. Confirm the configuration, terms and availability at the source.

Select up to three offers to compare their rates and terms.