Price floor Every GPU at least 30% below the market median — re-checked against the marketplace weekly.

See the proof

Guide · Costs & pricing

Cheapest cloud GPU for Stable Diffusion & Flux in 2026 ($/image)

What SDXL and Flux actually need, images-per-dollar on eight rentable GPUs, and where paying more per hour costs less per image.

9 min read Published 2026-08-27 Updated 2026-09-03 prices live from the sheet

Cheapest cloud GPU for Stable Diffusion & Flux in 2026 ($/image) — cover illustration

The answer table

Eight rentable cards, live prices, images per dollar (interruptible rate — image queues pause gracefully):

GPU$/hr int.SDXL s/imgSDXL img/$Flux s/imgFlux img/$
RTX 3060 12 GB $0.020 11.0 s 16,364 38 s 4,737
RTX 3090 24 GB $0.052 5.6 s 12,363 16 s 4,327
RTX 4070 12 GB $0.033 6.5 s 16,783 21 s 5,195
RTX 4090 24 GB $0.131 2.9 s 9,476 7.5 s 3,664
RTX 5090 32 GB $0.159 2.1 s 10,782 5.2 s 4,354
RTX A4000 16 GB $0.033 8.9 s 12,257 29 s 3,762
L40S 48 GB $0.233 3.4 s 4,544 9.0 s 1,717
RTX 5070 Ti 16 GB $0.056 4.8 s 13,393 15 s 4,286

Timings: community-typical SDXL 1024² @30 steps and Flux dev @20 steps; treat as ±20% and re-run on your workflow. Prices re-render live with every weekly sheet update.

Method: $/image, not $/hour

the whole method
images_per_dollar = 3600 / seconds_per_image / price_per_hour

This single division reorders the whole market. The "cheap" hourly cards at the top of the table lose to the RTX 4090 the moment throughput enters the equation — and the gap widens on Flux, where small cards spill to system RAM.

SDXL economics

  • Interactive sessions — an evening of prompting (~3 h) on a 4090: $0.79 on-demand. Use on-demand for sessions: an interruption mid-flow is worth more than the 50%.
  • Batches — queue overnight on interruptible; per the table, roughly 9,476 images per dollar.
  • Budget floor — the 3090 remains the best "always cheap" card: 24 GB means no workflow ever refuses to load.

Flux economics

  • 24 GB is the entry ticket (FP8 + offloaded text encoders). Below that, generation works but throughput collapses — the img/$ column shows the 3060's honest number.
  • 32 GB removes the ceiling — full-precision weights, LoRA stacks and upscalers resident: the 5090 leads img/$ despite the highest hourly rate on the table.
  • Serving users? Concurrency changes the criteria — see serverless endpoints with the ComfyUI template.

Three traps that triple bills

  1. Re-downloading models every session. 30 GB of checkpoints at every boot is slow and silly. A 120 GB volume costs $9.60/mo and mounts in seconds.
  2. Idle instances "kept for later". Stop them — a stopped instance bills only disk. The pause button is the biggest discount on this page.
  3. Benchmarking with someone else's workflow. Your LoRA count, resolution and sampler move s/image by 3×. Rent two candidates for one hour ($0.580 total) and measure your own pipeline.

Put the numbers to work

Every price in this guide is our live rate — fixed, ≥30% under the market median, billed per second. Deploy the exact setup above from the console in about 30 seconds, paid in crypto, no card and no KYC.

Deploy your first GPU in under a minute

Top up in crypto, benchmark us against your current provider. Per-second billing, fixed prices ≥ 30% below market — cancel by just stopping the instance.