Consumer flagship · Blackwell · launched 2025 · prices checked 2026-09-03
Rent NVIDIA RTX 5090 — 32 GB, $0.318/hr on-demand
- VRAM 32 GBGDDR7
- FP16 tensor 419TFLOPS
- PowerScore 295RTX 3090 = 100
- Configs 1–8×PCIe 5.0
- Online now 516 22 regions
Billed per second, price locked at deploy. Storage $0.08/GB/mo · bandwidth $0.01/GB — the whole fee schedule. Next weekly market re-check: 2026-09-10.
Consumer flagship · Blackwell architecture
The RTX 5090 is the Blackwell consumer flagship: 32 GB GDDR7 at 1.79 TB/s, 21,760 CUDA cores and native FP4 in fifth-generation tensor cores. It is the best dollars-per-token card on the sheet for 7B–32B models and the entry point for video diffusion; supply is the deepest in the catalogue.
The community favourite: consumer pricing with serious tensor throughput. Ideal for diffusion models, quantized LLMs and fine-tuning runs that fit in 32 GB. Supply is deep, so interruptible capacity is almost always available at half price.
NVIDIA RTX 5090 specs: VRAM, TFLOPS, bandwidth
| GPU model | NVIDIA RTX 5090 | Architecture | Blackwell (2025) |
|---|---|---|---|
| VRAM | 32 GB GDDR7 | Memory bandwidth | 1,792 GB/s |
| FP16 tensor perf. | 419 TFLOPS | FP32 perf. | 104.8 TFLOPS |
| CUDA cores | 21,760 | TDP | 575 W |
| PowerScore (RTX 3090 = 100) | 295 | PCIe generation | Gen 5.0 |
| Multi-GPU | 1× – 8× | Max instance storage | 4,000 GB NVMe |
| Network up to | 2,500 Mbps | CUDA | 12.4 – 13.0 |
Bandwidth, CUDA cores, TDP and FP32 are public NVIDIA figures; FP16 tensor is the dense (non-sparsity) number. Machine-level values come from live inventory.
RTX 5090 price per hour: on-demand, interruptible, reserved
One public rule sets every price on this page: the marketplace median for the RTX 5090 ($0.45/hr, snapshot 2026-09-03) × 0.70, rounded down — so on-demand is $0.318, 30% below market. Interruptible halves it; a 3-month reservation takes another 35% off.
| Mode | Per GPU-hour | Per day (24 h) | Per month (730 h) | What you get |
|---|---|---|---|---|
| On-demand | $0.318 | $7.63 | $232 | Guaranteed capacity, price locked at deploy, stop anytime |
| Interruptible | $0.159 | $3.82 | $116 | Flat −50%; may pause under capacity pressure, disk kept, auto-requeue |
| Reserved (3 months) | $0.206 | $4.94 | $150 | −35% on on-demand, rate locked for the term, capacity held |
Per GPU: an 8× machine costs exactly 8× — no multi-GPU premium. Estimate a full month with storage and bandwidth in the GPU cost calculator.
What you can run on a RTX 5090 (32 GB VRAM)
With 32 GB of GDDR7, a single card holds a ~13B-parameter LLM in FP16 or up to ~49B parameters quantized to 4-bit, with room for KV-cache at practical context lengths. Scale to 8× GPUs on one machine for bigger models or bigger batches — the per-GPU price stays $0.318.
- Recommended for llm inference — Best $/token in class
- Recommended for fine-tuning — QLoRA 70B on one GPU
- Recommended for image generation — Flux dev ~2 s/image on 4090
- Recommended for video generation — 32–141 GB VRAM on tap
- One-click template: ComfyUI on a RTX 5090
- One-click template: Ollama on a RTX 5090
- One-click template: Kohya's GUI on a RTX 5090
- Sizing help: LLM VRAM requirements guide
RTX 5090 availability by region
516 × RTX 5090 across 30 machines, live from inventory:
São Paulo
Amsterdam
London
Stockholm
Ashburn, VA
Tel Aviv
Milan
Vancouver
Dallas, TX
Paris
Sydney
Helsinki
- +10 more
RTX 5090 vs alternatives: price per TFLOP
| GPU | VRAM | FP16 | On-demand | $ / TFLOP-hr |
|---|---|---|---|---|
| RTX 5090 this card | 32 GB | 419 | $0.318 | $0.76‰ |
| RTX 5080 | 16 GB | 225 | $0.140 | $0.62‰ |
| RTX 5070 Ti | 16 GB | 176 | $0.112 | $0.64‰ |
| RTX 5070 | 12 GB | 123 | $0.089 | $0.72‰ |
| RTX 5060 Ti | 16 GB | 92 | $0.079 | $0.86‰ |
‰ = dollars per 1,000 TFLOP-hours of FP16 — a rough value-for-compute yardstick across cards.
- RTX 5090 vs RTX 4090 — specs, price per hour, which to rent
- RTX 5090 vs L40S — specs, price per hour, which to rent
- RTX 5090 vs A100 SXM4 — specs, price per hour, which to rent
- RTX PRO 6000 WS vs RTX 5090 — specs, price per hour, which to rent
- All GPU comparisons
Renting a RTX 5090: frequently asked questions
How much does it cost to rent an NVIDIA RTX 5090 per hour?
$0.318 per GPU-hour on-demand — a fixed price set at least 30% below the current market median of $0.45. Interruptible capacity costs $0.159/hr and a 3-month reservation $0.206/hr. Around $232/month if you keep one running non-stop, billed per second.
What can a RTX 5090 with 32 GB VRAM run?
In LLM terms, roughly a 13B-parameter model in FP16 or up to ~49B parameters 4-bit quantized on a single card, with context headroom. Multi-GPU instances (up to 8× on current inventory) multiply that; diffusion and rendering workloads fit comfortably at this VRAM class.
Is the RTX 5090 available to rent right now?
Yes — 516 GPUs across 30 machines in 22 regions are listed as we render this page. Configurations go from 1× to 8×. Deploy from the console and it is running in about 30 seconds.
How do I deploy a RTX 5090?
Create an account (email + password, no card, no KYC), top up in crypto, open the console, filter by RTX 5090, pick a machine and a template such as PyTorch, vLLM or ComfyUI. The same deploy is one command with the CLI: powergpu launch --gpu rtx-5090 --template pytorch.
Why is the RTX 5090 cheaper here than on GPU marketplaces?
We price from the public marketplace median and fix our on-demand rate at least 30% below it, rounded down. The price is re-checked weekly (next check 2026-09-10) and published — no auctions, no per-host roulette, no bidding.