Workstation · Ampere · launched 2021 · prices checked 2026-09-03
Rent NVIDIA RTX A4000 — 16 GB, $0.066/hr on-demand
- VRAM 16 GBGDDR6
- FP16 tensor 76TFLOPS
- PowerScore 54RTX 3090 = 100
- Configs 1–8×PCIe 4.0
- Online now 199 17 regions
Billed per second, price locked at deploy. Storage $0.08/GB/mo · bandwidth $0.01/GB — the whole fee schedule. Next weekly market re-check: 2026-09-10.
Workstation · Ampere architecture
The RTX A4000 is a single-slot, 140 W Ampere workstation card with 16 GB of ECC GDDR6. Hosts run many of them per machine, which keeps the price low — a good fit for 7B 4-bit inference, CI pipelines and classic ML at a fraction of a 3090's rate.
Workstation silicon with ECC memory and studio-certified drivers. Renderers (Blender, Octane, Redshift), CAD and simulation get the large VRAM they want without paying datacenter-flagship rates — and per-second billing suits render queues perfectly.
NVIDIA RTX A4000 specs: VRAM, TFLOPS, bandwidth
| GPU model | NVIDIA RTX A4000 | Architecture | Ampere (2021) |
|---|---|---|---|
| VRAM | 16 GB GDDR6 | Memory bandwidth | 448 GB/s |
| FP16 tensor perf. | 76 TFLOPS | FP32 perf. | 19.2 TFLOPS |
| CUDA cores | 6,144 | TDP | 140 W |
| PowerScore (RTX 3090 = 100) | 54 | PCIe generation | Gen 4.0 |
| Multi-GPU | 1× – 8× | Max instance storage | 8,000 GB NVMe |
| Network up to | 10,000 Mbps | CUDA | 12.4 – 13.0 |
Bandwidth, CUDA cores, TDP and FP32 are public NVIDIA figures; FP16 tensor is the dense (non-sparsity) number. Machine-level values come from live inventory.
RTX A4000 price per hour: on-demand, interruptible, reserved
One public rule sets every price on this page: the marketplace median for the RTX A4000 ($0.09/hr, snapshot 2026-09-03) × 0.70, rounded down — so on-demand is $0.066, 30% below market. Interruptible halves it; a 3-month reservation takes another 35% off.
| Mode | Per GPU-hour | Per day (24 h) | Per month (730 h) | What you get |
|---|---|---|---|---|
| On-demand | $0.066 | $1.58 | $48 | Guaranteed capacity, price locked at deploy, stop anytime |
| Interruptible | $0.033 | $0.79 | $24 | Flat −50%; may pause under capacity pressure, disk kept, auto-requeue |
| Reserved (3 months) | $0.042 | $1.01 | $31 | −35% on on-demand, rate locked for the term, capacity held |
Per GPU: an 8× machine costs exactly 8× — no multi-GPU premium. Estimate a full month with storage and bandwidth in the GPU cost calculator.
What you can run on a RTX A4000 (16 GB VRAM)
With 16 GB of GDDR6, a single card holds a ~3B-parameter LLM in FP16 or up to ~24B parameters quantized to 4-bit, with room for KV-cache at practical context lengths. Scale to 8× GPUs on one machine for bigger models or bigger batches — the per-GPU price stays $0.066.
- One-click template: Linux Desktop on a RTX A4000
- One-click template: ComfyUI on a RTX A4000
- One-click template: Ubuntu Desktop VM on a RTX A4000
- Sizing help: LLM VRAM requirements guide
RTX A4000 availability by region
199 × RTX A4000 across 30 machines, live from inventory:
Frankfurt
Chicago, IL
Amsterdam
Tokyo
New York, NY
Paris
Seattle, WA
London
Sydney
Los Angeles, CA
Mumbai
Ashburn, VA
- +5 more
RTX A4000 vs alternatives: price per TFLOP
| GPU | VRAM | FP16 | On-demand | $ / TFLOP-hr |
|---|---|---|---|---|
| RTX A4000 this card | 16 GB | 76 | $0.066 | $0.87‰ |
| Q RTX 6000 | 24 GB | 65 | $0.094 | $1.45‰ |
| Titan RTX | 24 GB | 65 | $0.094 | $1.45‰ |
| Quadro P4000 | 8 GB | 8 | $0.038 | $4.75‰ |
| RTX A2000 | 6 GB | 32 | $0.030 | $0.94‰ |
‰ = dollars per 1,000 TFLOP-hours of FP16 — a rough value-for-compute yardstick across cards.
Renting a RTX A4000: frequently asked questions
How much does it cost to rent an NVIDIA RTX A4000 per hour?
$0.066 per GPU-hour on-demand — a fixed price set at least 30% below the current market median of $0.09. Interruptible capacity costs $0.033/hr and a 3-month reservation $0.042/hr. Around $48/month if you keep one running non-stop, billed per second.
What can a RTX A4000 with 16 GB VRAM run?
In LLM terms, roughly a 3B-parameter model in FP16 or up to ~24B parameters 4-bit quantized on a single card, with context headroom. Multi-GPU instances (up to 8× on current inventory) multiply that; diffusion and rendering workloads fit comfortably at this VRAM class.
Is the RTX A4000 available to rent right now?
Yes — 199 GPUs across 30 machines in 17 regions are listed as we render this page. Configurations go from 1× to 8×. Deploy from the console and it is running in about 30 seconds.
How do I deploy a RTX A4000?
Create an account (email + password, no card, no KYC), top up in crypto, open the console, filter by RTX A4000, pick a machine and a template such as PyTorch, vLLM or ComfyUI. The same deploy is one command with the CLI: powergpu launch --gpu rtx-a4000 --template pytorch.
Why is the RTX A4000 cheaper here than on GPU marketplaces?
We price from the public marketplace median and fix our on-demand rate at least 30% below it, rounded down. The price is re-checked weekly (next check 2026-09-10) and published — no auctions, no per-host roulette, no bidding.