NVIDIA A100 40GB price per provider
NVIDIA A100 40GB price history
Loading the daily history…
Which models fit on one NVIDIA A100 40GB
With 4-bit weights and an 8K-token context, a single NVIDIA A100 40GB runs: Ornith 1.5 35B-A3B, Qwen3.6 35B-A3B, Qwen3 32B, Nemotron 3 Nano 30B-A3B, Gemma 4 31B, Qwen3-Coder 30B-A3B, JEV 27B VL, Qwen3.8 27B, Clef, Gemma 4 26B-A4B, Mistral Small 3.2 24B, LFM2 24B-A2B, gpt-oss 20B, Mellum2.1 12B-A2.5B, Gemma 4 12B, Humanizer, Ornith 1.5 9B, Qwen3.5 9B, Clef Flash, LFM2.5 8B-A1B, Llama 3.1 8B, Qwen3.5 4B, LightOnOCR-3 4B, Spark-X2.5 4B, D1 3B.
NVIDIA A100 40GB rental FAQ
How much does it cost to rent an NVIDIA A100 40GB?
On 2026-10-10, the cheapest on-demand NVIDIA A100 40GB costs $0.401 per GPU-hour on Vast.ai; across 3 providers prices range from $0.401 to $3.40. An 8-GPU server costs about $3.21 per hour, and one GPU running all month about $292.73.
Which cloud is cheapest for the NVIDIA A100 40GB?
Vast.ai at $0.401 per hour on 2026-10-10. Marketplace prices (Vast.ai) depend on the host and move within the day; the history on this page shows how they evolve.
Which LLMs can run on a single NVIDIA A100 40GB?
With 4-bit weights and an 8K-token context: Ornith 1.5 35B-A3B, Qwen3.6 35B-A3B, Qwen3 32B, Nemotron 3 Nano 30B-A3B, Gemma 4 31B, Qwen3-Coder 30B-A3B, JEV 27B VL, Qwen3.8 27B, Clef, Gemma 4 26B-A4B, Mistral Small 3.2 24B, LFM2 24B-A2B, gpt-oss 20B, Mellum2.1 12B-A2.5B, Gemma 4 12B, Humanizer, Ornith 1.5 9B, Qwen3.5 9B, Clef Flash, LFM2.5 8B-A1B, Llama 3.1 8B, Qwen3.5 4B, LightOnOCR-3 4B, Spark-X2.5 4B, D1 3B. Larger models need several GPUs; the LLM VRAM calculator sizes them.
Other GPUs
NVIDIA B300 / GB300 · NVIDIA B200 (HGX) · NVIDIA H200 · NVIDIA H100 NVL · NVIDIA H100 · NVIDIA A100 80GB · NVIDIA L40S · NVIDIA L4 · AMD MI355X / MI350X · AMD MI300X · RTX PRO 6000 Blackwell · RTX 6000 Ada · GeForce RTX 5090 · GeForce RTX 4090 · GeForce RTX 3090 · GeForce RTX 5080