Price floor Every GPU at least 30% below the market median — re-checked against the marketplace weekly.

See the proof

Compare · prices checked 2026-09-03

L40S vs H100 PCIE: specs, price per hour, which to rent

NVIDIA L40S (48 GB, $0.466/hr) against NVIDIA H100 PCIE (80 GB, $2.147/hr): public specs side by side, live fixed prices from our sheet, what fits in each card's VRAM, and a verdict written for real workloads — both are rentable right now.

L40S vs H100 PCIE specifications

Public NVIDIA figures (dense, non-sparsity). The last column is L40S relative to H100 PCIE.

SpecL40SH100 PCIEDifference
ArchitectureAda Lovelace (2023)Hopper (2022)
VRAM48 GB GDDR680 GB HBM2e−40%
Memory bandwidth864 GB/s2,000 GB/s−57%
FP16 tensor (dense)362 TFLOPS756 TFLOPS−52%
FP3291.6 TFLOPS51.0 TFLOPS+80%
CUDA cores18,17614,592+25%
TDP350 W350 Wsame
PCIe · NVLinkGen 4.0 · no NVLinkGen 5.0 · no NVLink
PowerScore (RTX 3090 = 100)255532−52%
Max GPUs per machine

L40S vs H100 PCIE price per hour

Fixed rates from our sheet — every on-demand price is the marketplace median × 0.70, rounded down. Monthly = 730 hours.

RateL40SH100 PCIECheaper
On-demand, per GPU-hour$0.466$2.147L40S (−78%)
Interruptible, per GPU-hour$0.233$1.073L40S
Reserved (3 mo), per GPU-hour$0.302$1.395L40S
On-demand, per month$340$1,567L40S
Market median (reference)$0.67$3.07
$ per 1,000 FP16 TFLOP-hours$1.29$2.84L40S (better value)
$ per GB of VRAM per hour$0.0097$0.0268L40S (better value)

Try a full month with storage and bandwidth in the GPU cost calculator.

What fits in VRAM: 48 GB vs 80 GB

WorkloadL40SH100 PCIE
Largest LLM in FP16, one card~14B~32B
Largest LLM at 4-bit, one card~72B~123B
Flux dev (FP8, ~17 GB)fitsfits
Wan 2.x 14B video (offloaded)yesyes
70B 4-bit LLM on one cardyesyes

Rules of thumb: ~2.4 GB per billion parameters in FP16 all-in, ~0.62 GB in 4-bit. Full tables in the VRAM guide.

Verdict: which should you rent?

The H100 PCIe has 80 GB of HBM2e, 2 TB/s and roughly double the tensor throughput; the L40S has 48 GB of GDDR6 and costs a lot less. The L40S is the value choice for models under 32B; the H100 PCIe for 70B and throughput-bound serving.

  • Cheaper per hour: L40S ($0.466 vs $2.147, −78%).
  • More VRAM: H100 PCIE (80 GB vs 48 GB).
  • More FP16 throughput: H100 PCIE (about 2.1×).
  • Best value per TFLOP-hour: L40S.
  • Best value per GB of VRAM: L40S.
  • Multi-GPU: L40S over PCIe · H100 PCIE over PCIe.
Is the H100 PCIE faster than the L40S?

On dense FP16 tensor throughput the H100 PCIE leads by about 2.1× (756 vs 362 TFLOPS). Memory bandwidth matters as much for inference: L40S 864 GB/s vs H100 PCIE 2,000 GB/s.

Which is cheaper to rent, the L40S or the H100 PCIE?

The L40S: $0.466/hr on-demand versus $2.147/hr — 78% less. Interruptible rates are $0.233 (L40S) and $1.073 (H100 PCIE). Per TFLOP-hour the better value is the L40S.

Which has more VRAM and what does that change?

The H100 PCIE has 80 GB versus 48 GB. In LLM terms that is roughly a 32B FP16 model (or ~123B in 4-bit) on one card against 14B FP16 (~72B 4-bit). If the model does not fit, speed is irrelevant.

Can I rent both on PowerGPU right now?

Yes — 59 × L40S and 60 × H100 PCIE are online as this page renders, deployable in about 30 seconds, billed per second, paid in crypto with no KYC.

Deploy your first GPU in under a minute

Top up in crypto, benchmark us against your current provider. Per-second billing, fixed prices ≥ 30% below market — cancel by just stopping the instance.