$GPU Rental Prices.com

LLM VRAM requirements

Model weights are only part of GPU memory use. Compare weight sizes below, then open a model to account for context, runtime overhead and training memory.

ModelParameters4-bit weights, GBCache estimate
Qwen2.5 0.5B0.49 billion0.2Config supported
Qwen2.5 Coder 0.5B0.49 billion0.2Config supported
Qwen3 0.6B0.75 billion0.4Config supported
gemma 3 1b1.00 billion0.5Not available
Llama 3.2 1B1.24 billion0.6Not available
Qwen2.5 1.5B1.54 billion0.8Config supported
Qwen2.5 Coder 1.5B1.54 billion0.8Config supported
DeepSeek R1 Distill Qwen 1.5B1.78 billion0.9Config supported
Qwen3 1.7B2.03 billion1.0Config supported
gemma 2 2b2.61 billion1.3Not available
Qwen2.5 3B3.09 billion1.5Config supported
Qwen2.5 Coder 3B3.09 billion1.5Config supported
Llama 3.2 3B3.21 billion1.6Not available
Phi 3 mini 4k3.82 billion1.9Not available
Phi 4 mini3.84 billion1.9Not available
Qwen3 4B4.02 billion2.0Config supported
gemma 3 4b4.30 billion2.2Not available
Mistral 7B v0.37.25 billion3.6Config supported
Qwen2.5 7B7.62 billion3.8Config supported
Qwen2.5 Coder 7B7.62 billion3.8Config supported
DeepSeek R1 Distill Qwen 7B7.62 billion3.8Config supported
Llama 3.1 8B8.03 billion4.0Not available
DeepSeek R1 Distill Llama 8B8.03 billion4.0Config supported
Qwen3 8B8.19 billion4.1Config supported
gemma 2 9b9.24 billion4.6Not available
gemma 3 12b12.19 billion6.1Not available
Mistral Nemo 240712.25 billion6.1Config supported
Phi 3 medium 4k13.96 billion7.0Not available
phi 414.66 billion7.3Config supported
Qwen3 14B14.77 billion7.4Config supported
Qwen2.5 14B14.77 billion7.4Config supported
Qwen2.5 Coder 14B14.77 billion7.4Config supported
DeepSeek R1 Distill Qwen 14B14.77 billion7.4Config supported
Mistral Small 24B 250123.57 billion11.8Config supported
gemma 2 27b27.23 billion13.6Not available
gemma 3 27b27.43 billion13.7Not available
Qwen3 30B A3B30.53 billion15.3Config supported
Qwen3 32B32.76 billion16.4Config supported
Qwen2.5 32B32.76 billion16.4Config supported
Qwen2.5 Coder 32B32.76 billion16.4Config supported
DeepSeek R1 Distill Qwen 32B32.76 billion16.4Config supported
Mixtral 8x7B v0.146.70 billion23.4Config supported
Llama 3.1 70B70.55 billion35.3Not available
Llama 3.3 70B70.55 billion35.3Not available
DeepSeek R1 Distill Llama 70B70.55 billion35.3Config supported
Qwen2.5 72B72.71 billion36.4Config supported
Mixtral 8x22B v0.1140.63 billion70.3Config supported
Qwen3 235B A22B235.09 billion117.5Config supported
Llama 3.1 405B405.85 billion202.9Not available

Estimate training cost · Compare LLM API pricing · GPU rental prices