Cross-posted from Best GPU for AI — visit the original for our VRAM calculator, GPU comparison table, and current Amazon pricing.
Quick answer: The cheapest cloud GPU that can actually do modern AI work is a rented RTX 4090 at roughly $0.35/hr on marketplace platforms like Vast.ai (as of mid-2026). You can rent GPUs for as little as $0.02/hr, but those bottom-tier cards are so slow on current models that they usually cost you more per finished job. Cheapest sticker price and cheapest way to get work done are two different questions.
That distinction is the whole game with budget cloud compute. A $0.02/hr RTX 3060 sounds unbeatable until you realize your Flux render queue takes eight times longer on it. This guide ranks the cheap options and does the math on which are genuine and which are false economy.
Cloud GPU prices ranked, cheapest first
These are marketplace floor prices as of mid-2026 — prices move weekly, so treat them as anchors, not quotes:
| GPU | From (on-demand) | VRAM | Honest verdict |
|---|---|---|---|
| RTX 3060 12GB | ~$0.02/hr | 12GB | Dirt cheap, painfully slow for modern image models |
| RTX 3090 | ~$0.15-0.25/hr | 24GB | Solid budget pick for 13B LLMs |
| RTX 4090 | ~$0.35/hr | 24GB | The value king — best $/task in cloud AI |
| RTX 5000-class | ~$0.39/hr | 24-32GB | Newer, marginally faster, similar value |
| A100 80GB | ~$0.75-1.50/hr | 80GB | Cheapest serious training card |
| H100 | ~$2.00/hr | 80GB | Only when you genuinely need the speed |
Two footnotes that matter more than the table. First, spot/interruptible pricing runs roughly 50-80% below on-demand — but the host can reclaim your machine mid-job. Second, hyperscalers (AWS, GCP, Azure) charge roughly 2-3x more than GPU-first providers for the same silicon. If you're hunting cheapest, skip the hyperscalers.
For the full landscape of platforms, plans, and setup, our GPU rental for AI guide is the companion piece to this one.
The $/task math: why $0.02/hr can be the expensive option
Here's the trap. Say you're generating a batch of 500 SDXL images. On a rented RTX 4090 at ~$0.35/hr, that batch might take around an hour — call it $0.35 all-in. On a $0.02/hr RTX 3060, the same batch could easily take 8-10 hours. That's about $0.20 in rental fees — the slow card technically wins on paper.
Except it doesn't, for three reasons. You're paying storage the whole time the instance exists. You're babysitting a job overnight instead of iterating (your time is not free, even hobbyist time). And the moment you step up to anything heavier — Flux, video models, a LoRA training run — the slow card stops merely losing and starts failing outright.
LoRA training makes the gap brutal. A fine-tune that finishes in roughly 2 hours on a 4090 (~$0.70) can grind for 12-15 hours on a bottom-tier card — if the older architecture can run modern training recipes efficiently at all. The $/hr number on the pricing page is an input. The number that matters is $/finished-job, and on that metric the rented 4090 is the best deal in AI compute right now.
What the ultra-cheap tiers can't do at all
Some workloads aren't slow on cheap cards — they're effectively impossible:
- Modern video models (Wan, Hunyuan-class): need 24GB+ and serious compute. A 12GB card at any price is a non-starter.
- Flux at usable speed: it technically runs quantized on 12GB, but generation times stretch to the point where iteration dies.
- 34B+ LLMs: 12GB doesn't fit them at any sane quantization. Even 24GB is the floor, and 48GB+ is where they get comfortable.
- Full fine-tuning of anything beyond toy models: this is A100 territory, minimum.
The honest framing: the $0.02-0.10/hr tier is for learning, running small 7B models, and Stable Diffusion 1.5. That's a legitimate use case — it's just not "doing modern AI work cheaply." It's doing 2023 AI work cheaply.
The hidden fees nobody mentions
The $/hr rate isn't your bill. Watch for:
- Storage: persistent volumes bill 24/7 whether the GPU runs or not. A 200GB volume quietly adds up while you sleep.
- Egress: downloading your outputs and checkpoints typically costs $0.05-0.12/GB on many platforms. Pull down 100GB of training checkpoints and you've paid more for the download than the training.
- Minimum balances and credits: several platforms make you pre-load $10-25 before you can rent anything, which matters when your actual compute bill is fifty cents.
- Idle instances: stopped is not terminated — stopped instances often keep billing for storage.
None are scandalous individually; together they routinely double a casual user's effective rate.
Which cheap cloud GPU should you rent?
- Just learning, SD 1.5, small 7B models? Grab the cheapest RTX 3060/3090-class listing you can find — slowness doesn't hurt at this scale.
- SDXL, Flux, 13B LLMs, LoRA training? Rent a 4090 at ~$0.35/hr. Full stop. It's the cheapest card that's actually fast on these workloads.
- Fine-tuning bigger models or need 80GB? Spot-priced A100s (often under $1/hr interruptible) are the budget path — checkpoint aggressively.
- Renting 20+ hours/month, every month? Do the ownership math. Our cloud GPU vs home GPU breakdown covers the break-even point — it arrives faster than you'd expect.
If your workload is light enough that a $0.02/hr rental covers it, a used local card covers it too — permanently, with no egress fees and no clock running. (The quiet third option most cloud pricing articles skip.)
See the recommended pick on the original guide
Common mistakes that burn budget renters
- Leaving instances running. The classic. A forgotten 4090 at $0.35/hr is $8+/day; a forgotten A100 is a genuinely bad week. Set spend alerts on day one.
- Picking by $/hr alone. As above — a card that's 5x cheaper but 10x slower loses on $/task, every time.
- Running long training on spot without checkpoints. Interruptible instances get reclaimed. If you haven't saved state, that 50-80% discount just bought you a full restart.
- Ignoring the host's bandwidth on marketplace platforms. A cheap listing with slow internet means model downloads eat billed time before you compute anything. Vast.ai shows bandwidth per listing — filter for it.
Final verdict
| If you want... | Rent this | Expect (mid-2026) |
|---|---|---|
| Absolute lowest $/hr | RTX 3060 12GB | ~$0.02-0.05/hr, slow |
| Best $/task (most people) | RTX 4090 | ~$0.35/hr on-demand |
| Cheapest 80GB for training | A100 80GB spot | often under $1/hr, interruptible |
| Max speed, money no object | H100 | ~$2.00/hr and up |
Marketplace platforms are where these floor prices live — RunPod vs Vast.ai compares the two biggest head-to-head. If you'd rather own a card outright, our best GPU for AI under $500 roundup covers the buy-side of the same budget question.
Cheap cloud GPU questions, answered
What is the cheapest cloud GPU?
The cheapest listings are RTX 3060-class cards from roughly $0.02/hr on marketplace platforms like Vast.ai, as of mid-2026. But the cheapest GPU worth renting for modern AI work is an RTX 4090 at roughly $0.35/hr — slower cards take so much longer per job that they often cost more per finished task despite the lower hourly rate.
Is a cheap cloud GPU good enough for Stable Diffusion?
For Stable Diffusion 1.5, yes — even a sub-$0.05/hr card handles it fine. For SDXL and Flux, cheap 12GB cards run them slowly enough that iteration becomes painful. If you're doing serious image generation in 2026, a rented RTX 4090 at roughly $0.35/hr is the practical floor.
Why is Vast.ai so cheap?
Vast.ai is a marketplace where individual hosts and small datacenters rent out their own hardware, so prices are set by open competition rather than corporate rate cards. That makes it roughly 50-80% cheaper than hyperscalers for the same GPU. The trade-off is variable reliability — host quality, bandwidth, and uptime differ per listing.
Is spot pricing safe for training?
Spot or interruptible instances are roughly 50-80% cheaper, but the provider can reclaim the machine with little warning. They're safe for training only if you checkpoint frequently and your framework can resume from the last save. For short jobs and inference they're excellent; for long uncheckpointed runs they're a gamble.
If a cloud GPU is too slow to finish your job today, it was never cheap — rent the ~$0.35/hr 4090 and pay for results, not hours.
Related guides on Best GPU for AI
- GPU Rental for AI: What to Rent and What It Costs (2026)
- RunPod vs Vast.ai for AI Workloads in 2026 (Compared)
- Best Budget GPU for AI in 2026 (5 Picks From $150)
Read the full guide on Best GPU for AI — includes our VRAM calculator, GPU comparison table, and live pricing.
Top comments (0)