Datacenter · Ada Lovelace · launched 2023 · prices checked 2026-09-03
Rent NVIDIA L4 — 24 GB, $0.225/hr on-demand
- VRAM 24 GBGDDR6
- FP16 tensor 121TFLOPS
- PowerScore 85RTX 3090 = 100
- Configs 1–4×PCIe 4.0
- Online now 24 4 regions
Billed per second, price locked at deploy. Storage $0.08/GB/mo · bandwidth $0.01/GB — the whole fee schedule. Next weekly market re-check: 2026-09-10.
Datacenter · Ada Lovelace architecture
The L4 is the efficient inference card: 24 GB GDDR6 at just 72 W, single-slot, Ada tensor cores and dual NVENC/NVDEC engines. It is the right rental for video transcoding pipelines, small-model APIs and batch embeddings where power and price matter more than peak FLOPS.
A server-class card built for sustained 24/7 load: passive cooling in proper chassis, ECC memory, and drivers validated for compute. The sweet spot for inference fleets and fine-tuning jobs that need stability more than headline FLOPS.
NVIDIA L4 specs: VRAM, TFLOPS, bandwidth
| GPU model | NVIDIA L4 | Architecture | Ada Lovelace (2023) |
|---|---|---|---|
| VRAM | 24 GB GDDR6 | Memory bandwidth | 300 GB/s |
| FP16 tensor perf. | 121 TFLOPS | FP32 perf. | 30.3 TFLOPS |
| CUDA cores | 7,680 | TDP | 72 W |
| PowerScore (RTX 3090 = 100) | 85 | PCIe generation | Gen 4.0 |
| Multi-GPU | 1× – 4× | Max instance storage | 8,000 GB NVMe |
| Network up to | 10,000 Mbps | CUDA | 12.4 – 13.0 |
Bandwidth, CUDA cores, TDP and FP32 are public NVIDIA figures; FP16 tensor is the dense (non-sparsity) number. Machine-level values come from live inventory.
L4 price per hour: on-demand, interruptible, reserved
One public rule sets every price on this page: the marketplace median for the L4 ($0.32/hr, snapshot 2026-09-03) × 0.70, rounded down — so on-demand is $0.225, 30% below market. Interruptible halves it; a 3-month reservation takes another 35% off.
| Mode | Per GPU-hour | Per day (24 h) | Per month (730 h) | What you get |
|---|---|---|---|---|
| On-demand | $0.225 | $5.40 | $164 | Guaranteed capacity, price locked at deploy, stop anytime |
| Interruptible | $0.112 | $2.69 | $82 | Flat −50%; may pause under capacity pressure, disk kept, auto-requeue |
| Reserved (3 months) | $0.146 | $3.50 | $107 | −35% on on-demand, rate locked for the term, capacity held |
Per GPU: an 4× machine costs exactly 4× — no multi-GPU premium. Estimate a full month with storage and bandwidth in the GPU cost calculator.
What you can run on a L4 (24 GB VRAM)
With 24 GB of GDDR6, a single card holds a ~8B-parameter LLM in FP16 or up to ~32B parameters quantized to 4-bit, with room for KV-cache at practical context lengths. Scale to 4× GPUs on one machine for bigger models or bigger batches — the per-GPU price stays $0.225.
- One-click template: vLLM on a L4
- One-click template: PyTorch on a L4
- One-click template: ComfyUI on a L4
- Sizing help: LLM VRAM requirements guide
L4 availability by region
24 × L4 across 4 machines, live from inventory:
São Paulo
Montréal
Amsterdam
Frankfurt
L4 vs alternatives: price per TFLOP
| GPU | VRAM | FP16 | On-demand | $ / TFLOP-hr |
|---|---|---|---|---|
| L4 this card | 24 GB | 121 | $0.225 | $1.86‰ |
| L40 | 48 GB | 181 | $0.234 | $1.29‰ |
| A10 | 24 GB | 125 | $0.168 | $1.34‰ |
| Tesla T4 | 16 GB | 65 | $0.104 | $1.60‰ |
| Tesla V100 | 16 GB | 125 | $0.089 | $0.71‰ |
‰ = dollars per 1,000 TFLOP-hours of FP16 — a rough value-for-compute yardstick across cards.
- L4 vs Tesla T4 — specs, price per hour, which to rent
- L4 vs A10 — specs, price per hour, which to rent
- All GPU comparisons
Renting a L4: frequently asked questions
How much does it cost to rent an NVIDIA L4 per hour?
$0.225 per GPU-hour on-demand — a fixed price set at least 30% below the current market median of $0.32. Interruptible capacity costs $0.112/hr and a 3-month reservation $0.146/hr. Around $164/month if you keep one running non-stop, billed per second.
What can a L4 with 24 GB VRAM run?
In LLM terms, roughly a 8B-parameter model in FP16 or up to ~32B parameters 4-bit quantized on a single card, with context headroom. Multi-GPU instances (up to 4× on current inventory) multiply that; diffusion and rendering workloads fit comfortably at this VRAM class.
Is the L4 available to rent right now?
Yes — 24 GPUs across 4 machines in 4 regions are listed as we render this page. Configurations go from 1× to 4×. Deploy from the console and it is running in about 30 seconds.
How do I deploy a L4?
Create an account (email + password, no card, no KYC), top up in crypto, open the console, filter by L4, pick a machine and a template such as PyTorch, vLLM or ComfyUI. The same deploy is one command with the CLI: powergpu launch --gpu l4 --template pytorch.
Why is the L4 cheaper here than on GPU marketplaces?
We price from the public marketplace median and fix our on-demand rate at least 30% below it, rounded down. The price is re-checked weekly (next check 2026-09-10) and published — no auctions, no per-host roulette, no bidding.