Live · refreshed hourly
Cloud GPU prices right now
What it actually costs to rent each GPU this hour — and, further down, the cheapest card that genuinely runs each open model. Real offers from the providers we track, not a marketing rate card.
New here? Start with these
- Can you run your own AI without buying a graphics card? (yes — here's how) →
- What is a cloud GPU? (explained like you're completely new) →
- Do you need your own GPU to build an AI app, or can you just use an API? →
- Want to make AI images but don't have a good GPU? Your two options, honestly →
- Do you actually need a GPU to use AI for your business? (almost certainly not) →
Cheapest NVIDIA GeForce RTX 5090 right now
$0.43/hr
· spot · 32 GB · ≈ $103/mo at 8h/day
Every GPU we price — cheapest first
Minimum live $/hr per provider (median in grey when it differs). Spot is interruptible; on-demand is reserved while you run it.
| GPU | VRAM | Vast.ai · spot | Runpod · on-demand | Runpod · spot | Cheapest | Rent |
|---|---|---|---|---|---|---|
| NVIDIA Tesla V100 32GB SXM2 | 32 GB | $0.03 · med $0.19 | $0.19 | $0.19 | $0.03/hr | |
| NVIDIA GeForce RTX 3060 12GB | 12 GB | $0.04 · med $0.07 | — | — | $0.04/hr | |
| NVIDIA RTX A4000 | 16 GB | $0.07 · med $0.11 | $0.17 | $0.17 | $0.07/hr | |
| NVIDIA GeForce RTX 4060 Ti 16GB | 16 GB | $0.08 | — | — | $0.08/hr | |
| NVIDIA GeForce RTX 5060 Ti 16GB | 16 GB | $0.09 · med $0.10 | — | — | $0.09/hr | |
| NVIDIA GeForce RTX 4070 SUPER | 12 GB | $0.09 · med $0.16 | — | — | $0.09/hr | |
| NVIDIA GeForce RTX 4070 Ti SUPER | 16 GB | $0.09 · med $0.10 | $0.19 | $0.19 | $0.09/hr | |
| NVIDIA L40S | 48 GB | $0.11 · med $0.26 | $0.44 | $0.44 | $0.11/hr | |
| NVIDIA GeForce RTX 3090 Ti | 24 GB | $0.12 · med $0.17 | $0.22 | $0.22 | $0.12/hr | |
| NVIDIA GeForce RTX 3080 12GB | 12 GB | $0.13 · med $0.20 | $0.17 | $0.17 | $0.13/hr | |
| NVIDIA GeForce RTX 4070 | 12 GB | $0.13 · med $0.25 | — | — | $0.13/hr | |
| NVIDIA GeForce RTX 4090 | 24 GB | $0.14 · med $0.53 | $0.34 | $0.34 | $0.14/hr | |
| NVIDIA Tesla T4 | 16 GB | $0.15 · med $0.15 | — | — | $0.15/hr | |
| NVIDIA RTX A5000 | 24 GB | $0.20 · med $0.23 | $0.16 | $0.16 | $0.16/hr | |
| NVIDIA GeForce RTX 5070 Ti | 16 GB | $0.16 · med $0.28 | — | — | $0.16/hr | |
| NVIDIA GeForce RTX 3080 Ti | 12 GB | $0.16 · med $0.18 | $0.18 | $0.18 | $0.16/hr | |
| NVIDIA A100 80GB PCIe | 80 GB | $0.24 | — | — | $0.24/hr | |
| NVIDIA GeForce RTX 4080 SUPER | 16 GB | $0.27 | $0.27 | $0.27 | $0.27/hr | |
| NVIDIA GeForce RTX 5080 | 16 GB | $0.27 · med $0.33 | $0.39 | $0.39 | $0.27/hr | |
| NVIDIA GeForce RTX 4080 | 16 GB | $0.29 | — | — | $0.29/hr | |
| NVIDIA RTX A6000 | 48 GB | $0.40 · med $0.48 | $0.33 | $0.33 | $0.33/hr | |
| NVIDIA GeForce RTX 5090 | 32 GB | $0.43 · med $0.66 | $0.69 | $0.69 | $0.43/hr | |
| NVIDIA RTX 5000 Ada Generation | 32 GB | $0.47 | $0.49 | $0.49 | $0.47/hr | |
| AMD Instinct MI300X | 192 GB | — | $0.50 | $0.50 | $0.50/hr | |
| NVIDIA RTX 6000 Ada Generation | 48 GB | $0.54 · med $0.70 | $0.74 | $0.74 | $0.54/hr | |
| NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition | 96 GB | — | $1.69 | $1.69 | $1.69/hr | |
| NVIDIA H100 SXM | 80 GB | $1.74 · med $2.40 | $2.69 | $2.69 | $1.74/hr | |
| NVIDIA H100 PCIe | 80 GB | $2.00 · med $2.48 | $1.99 | $1.99 | $1.99/hr | |
| NVIDIA H200 SXM | 141 GB | $2.63 · med $4.74 | $3.59 | $3.59 | $2.63/hr | |
| NVIDIA B200 | 192 GB | $9.38 | $5.98 | $5.98 | $5.98/hr |
Cheapest way to run each model in the cloud
The cheapest live-priced GPU each model fits on at 16k context, computed from the model's own config — and what a month of it costs.
| Model | Cheapest GPU that fits | Quant | $/hr | 8h/day · mo | 24/7 · mo |
|---|---|---|---|---|---|
| gpt-oss 120B (MoE) | NVIDIA L40S · 48 GB | Q2_K47 GB | $0.11 | $26 | $77 |
| GLM-4.5-Air (MoE) | NVIDIA L40S · 48 GB | Q2_K45 GB | $0.11 | $26 | $77 |
| Qwen3-Next 80B-A3B (MoE) | NVIDIA L40S · 48 GB | Q4_K_M47 GB | $0.11 | $26 | $77 |
| Mixtral 8x7B (MoE) | NVIDIA Tesla V100 32GB SXM2 · 32 GB | Q4_K_M29 GB | $0.03 | $7 | $20 |
| Qwen2.5 Coder 32B | NVIDIA Tesla V100 32GB SXM2 · 32 GB | Q6_K30 GB | $0.03 | $7 | $20 |
| DeepSeek R1 Distill 32B | NVIDIA Tesla V100 32GB SXM2 · 32 GB | Q6_K30 GB | $0.03 | $7 | $20 |
| Qwen3 32B | NVIDIA Tesla V100 32GB SXM2 · 32 GB | Q6_K30 GB | $0.03 | $7 | $20 |
| Qwen3 30B-A3B (MoE) | NVIDIA Tesla V100 32GB SXM2 · 32 GB | Q6_K26 GB | $0.03 | $7 | $20 |
| Qwen3-Coder 30B-A3B (MoE) | NVIDIA Tesla V100 32GB SXM2 · 32 GB | Q6_K26 GB | $0.03 | $7 | $20 |
| Mistral Small 24B | NVIDIA Tesla V100 32GB SXM2 · 32 GB | Q8_027 GB | $0.03 | $7 | $20 |
| gpt-oss 20B (MoE) | NVIDIA Tesla V100 32GB SXM2 · 32 GB | Q8_022 GB | $0.03 | $7 | $20 |
| Qwen3 14B | NVIDIA Tesla V100 32GB SXM2 · 32 GB | F1631 GB | $0.03 | $7 | $20 |
| Phi-4 14B | NVIDIA Tesla V100 32GB SXM2 · 32 GB | F1631 GB | $0.03 | $7 | $20 |
| Qwen3 8B | NVIDIA Tesla V100 32GB SXM2 · 32 GB | F1618 GB | $0.03 | $7 | $20 |
| Qwen2.5 7B Instruct | NVIDIA Tesla V100 32GB SXM2 · 32 GB | F1616 GB | $0.03 | $7 | $20 |
| Mistral 7B Instruct v0.3 | NVIDIA Tesla V100 32GB SXM2 · 32 GB | F1616 GB | $0.03 | $7 | $20 |
Monthly figures are the live $/hr × 240h (8h/day) or × 720h (24/7). Whether renting beats buying depends on your own usage — the rent-vs-buy break-even walks through it honestly.
Get a note when H100 or 5090 rentals drop
Prices have fallen 60–75% from their 2024 peak and still move week to week. We watch them hourly; you get the drops that matter, not the noise.
The Aliteq brief
The tech worth knowing — hardware, AI, gaming, deals. No spam, unsubscribe anytime.
The cheapest spot marketplace
A marketplace of independent hosts — usually the lowest $/hr anywhere, with reliability that varies by host. Best for interruptible work.
Managed pods + serverless
Predictable on-demand pods, one-click templates and per-second serverless — the easiest way to go from zero to a running GPU.
How to read this page
How often are these cloud GPU prices updated?
Hourly. Our worker pulls live offers from each provider's API and stores the minimum, 25th-percentile and median $/hr per GPU. The capture time is shown on the page — never quote a price without its timestamp.
Which cloud GPU providers are tracked?
2 today: Vast.ai (a spot marketplace of independent hosts — cheapest, variable reliability) and Runpod (managed on-demand pods and spot). We add providers only where a live, legally usable price API exists; we don't scrape.
What's the difference between spot and on-demand GPU pricing?
Spot (interruptible) capacity is cheaper but can be reclaimed by the provider mid-job — fine for inference and checkpointed training. On-demand is reserved while you run it and costs more. The table shows both where a provider offers them.
How is 'cheapest GPU that runs a model' calculated?
From each model's own config (layers, hidden size, heads) we compute the VRAM it needs at each quantisation at 16k context (bits-per-weight constants validated to a 0.7% median error), then pick the cheapest live-priced GPU it fits on. It is an estimate of fit and cost, not a benchmark.
Disclosure:provider links may be referral links — if you sign up through one we may earn a commission at no cost to you. That never changes the ranking: every price on this page is the live figure from the provider's API, cheapest first. Sponsored placements are labelled. Provider names and logos are their owners' trademarks, used to identify the services compared; no partnership or endorsement is implied. Prices are pulled hourly; VRAM fit is an estimate from each model's config, not a benchmark.