ALITEQ.

Live · refreshed hourly

Cloud GPU prices right now

What it actually costs to rent each GPU this hour — and, further down, the cheapest card that genuinely runs each open model. Real offers from the providers we track, not a marketing rate card.

30 GPUs priced190 live offers2 providers trackedcaptured

New here? Start with these

Cheapest NVIDIA H100 SXM right now

$1.74/hr

Vast.ai· spot · 80 GB · ≈ $416/mo at 8h/day

Cheapest NVIDIA GeForce RTX 5090 right now

$0.43/hr

Vast.ai· spot · 32 GB · ≈ $103/mo at 8h/day

Every GPU we price — cheapest first

Minimum live $/hr per provider (median in grey when it differs). Spot is interruptible; on-demand is reserved while you run it.

GPUVRAMVast.ai · spotRunpod · on-demandRunpod · spotCheapestRent
NVIDIA Tesla V100 32GB SXM232 GB$0.03 · med $0.19$0.19$0.19$0.03/hrVast.aiVast.ai
NVIDIA GeForce RTX 3060 12GB12 GB$0.04 · med $0.07$0.04/hrVast.aiVast.ai
NVIDIA RTX A400016 GB$0.07 · med $0.11$0.17$0.17$0.07/hrVast.aiVast.ai
NVIDIA GeForce RTX 4060 Ti 16GB16 GB$0.08$0.08/hrVast.aiVast.ai
NVIDIA GeForce RTX 5060 Ti 16GB16 GB$0.09 · med $0.10$0.09/hrVast.aiVast.ai
NVIDIA GeForce RTX 4070 SUPER12 GB$0.09 · med $0.16$0.09/hrVast.aiVast.ai
NVIDIA GeForce RTX 4070 Ti SUPER16 GB$0.09 · med $0.10$0.19$0.19$0.09/hrVast.aiVast.ai
NVIDIA L40S48 GB$0.11 · med $0.26$0.44$0.44$0.11/hrVast.aiVast.ai
NVIDIA GeForce RTX 3090 Ti24 GB$0.12 · med $0.17$0.22$0.22$0.12/hrVast.aiVast.ai
NVIDIA GeForce RTX 3080 12GB12 GB$0.13 · med $0.20$0.17$0.17$0.13/hrVast.aiVast.ai
NVIDIA GeForce RTX 407012 GB$0.13 · med $0.25$0.13/hrVast.aiVast.ai
NVIDIA GeForce RTX 409024 GB$0.14 · med $0.53$0.34$0.34$0.14/hrVast.aiVast.ai
NVIDIA Tesla T416 GB$0.15 · med $0.15$0.15/hrVast.aiVast.ai
NVIDIA RTX A500024 GB$0.20 · med $0.23$0.16$0.16$0.16/hrRunpodRunpod
NVIDIA GeForce RTX 5070 Ti16 GB$0.16 · med $0.28$0.16/hrVast.aiVast.ai
NVIDIA GeForce RTX 3080 Ti12 GB$0.16 · med $0.18$0.18$0.18$0.16/hrVast.aiVast.ai
NVIDIA A100 80GB PCIe80 GB$0.24$0.24/hrVast.aiVast.ai
NVIDIA GeForce RTX 4080 SUPER16 GB$0.27$0.27$0.27$0.27/hrVast.aiVast.ai
NVIDIA GeForce RTX 508016 GB$0.27 · med $0.33$0.39$0.39$0.27/hrVast.aiVast.ai
NVIDIA GeForce RTX 408016 GB$0.29$0.29/hrVast.aiVast.ai
NVIDIA RTX A600048 GB$0.40 · med $0.48$0.33$0.33$0.33/hrRunpodRunpod
NVIDIA GeForce RTX 509032 GB$0.43 · med $0.66$0.69$0.69$0.43/hrVast.aiVast.ai
NVIDIA RTX 5000 Ada Generation32 GB$0.47$0.49$0.49$0.47/hrVast.aiVast.ai
AMD Instinct MI300X192 GB$0.50$0.50$0.50/hrRunpodRunpod
NVIDIA RTX 6000 Ada Generation48 GB$0.54 · med $0.70$0.74$0.74$0.54/hrVast.aiVast.ai
NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition96 GB$1.69$1.69$1.69/hrRunpodRunpod
NVIDIA H100 SXM80 GB$1.74 · med $2.40$2.69$2.69$1.74/hrVast.aiVast.ai
NVIDIA H100 PCIe80 GB$2.00 · med $2.48$1.99$1.99$1.99/hrRunpodRunpod
NVIDIA H200 SXM141 GB$2.63 · med $4.74$3.59$3.59$2.63/hrVast.aiVast.ai
NVIDIA B200192 GB$9.38$5.98$5.98$5.98/hrRunpodRunpod

Cheapest way to run each model in the cloud

The cheapest live-priced GPU each model fits on at 16k context, computed from the model's own config — and what a month of it costs.

ModelCheapest GPU that fitsQuant$/hr8h/day · mo24/7 · mo
gpt-oss 120B (MoE)NVIDIA L40S · 48 GBVast.aiVast.aiQ2_K47 GB$0.11$26$77
GLM-4.5-Air (MoE)NVIDIA L40S · 48 GBVast.aiVast.aiQ2_K45 GB$0.11$26$77
Qwen3-Next 80B-A3B (MoE)NVIDIA L40S · 48 GBVast.aiVast.aiQ4_K_M47 GB$0.11$26$77
Mixtral 8x7B (MoE)NVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiQ4_K_M29 GB$0.03$7$20
Qwen2.5 Coder 32BNVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiQ6_K30 GB$0.03$7$20
DeepSeek R1 Distill 32BNVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiQ6_K30 GB$0.03$7$20
Qwen3 32BNVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiQ6_K30 GB$0.03$7$20
Qwen3 30B-A3B (MoE)NVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiQ6_K26 GB$0.03$7$20
Qwen3-Coder 30B-A3B (MoE)NVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiQ6_K26 GB$0.03$7$20
Mistral Small 24BNVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiQ8_027 GB$0.03$7$20
gpt-oss 20B (MoE)NVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiQ8_022 GB$0.03$7$20
Qwen3 14BNVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiF1631 GB$0.03$7$20
Phi-4 14BNVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiF1631 GB$0.03$7$20
Qwen3 8BNVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiF1618 GB$0.03$7$20
Qwen2.5 7B InstructNVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiF1616 GB$0.03$7$20
Mistral 7B Instruct v0.3NVIDIA Tesla V100 32GB SXM2 · 32 GBVast.aiVast.aiF1616 GB$0.03$7$20

Monthly figures are the live $/hr × 240h (8h/day) or × 720h (24/7). Whether renting beats buying depends on your own usage — the rent-vs-buy break-even walks through it honestly.

Get a note when H100 or 5090 rentals drop

Prices have fallen 60–75% from their 2024 peak and still move week to week. We watch them hourly; you get the drops that matter, not the noise.

The Aliteq brief

The tech worth knowing — hardware, AI, gaming, deals. No spam, unsubscribe anytime.

Vast.aiReferral link

The cheapest spot marketplace

A marketplace of independent hosts — usually the lowest $/hr anywhere, with reliability that varies by host. Best for interruptible work.

Runpod

Managed pods + serverless

Predictable on-demand pods, one-click templates and per-second serverless — the easiest way to go from zero to a running GPU.

How to read this page

How often are these cloud GPU prices updated?

Hourly. Our worker pulls live offers from each provider's API and stores the minimum, 25th-percentile and median $/hr per GPU. The capture time is shown on the page — never quote a price without its timestamp.

Which cloud GPU providers are tracked?

2 today: Vast.ai (a spot marketplace of independent hosts — cheapest, variable reliability) and Runpod (managed on-demand pods and spot). We add providers only where a live, legally usable price API exists; we don't scrape.

What's the difference between spot and on-demand GPU pricing?

Spot (interruptible) capacity is cheaper but can be reclaimed by the provider mid-job — fine for inference and checkpointed training. On-demand is reserved while you run it and costs more. The table shows both where a provider offers them.

How is 'cheapest GPU that runs a model' calculated?

From each model's own config (layers, hidden size, heads) we compute the VRAM it needs at each quantisation at 16k context (bits-per-weight constants validated to a 0.7% median error), then pick the cheapest live-priced GPU it fits on. It is an estimate of fit and cost, not a benchmark.

Disclosure:provider links may be referral links — if you sign up through one we may earn a commission at no cost to you. That never changes the ranking: every price on this page is the live figure from the provider's API, cheapest first. Sponsored placements are labelled. Provider names and logos are their owners' trademarks, used to identify the services compared; no partnership or endorsement is implied. Prices are pulled hourly; VRAM fit is an estimate from each model's config, not a benchmark.