Price floor Every GPU at least 30% below the market median — re-checked against the marketplace weekly.

See the proof

Compare · prices checked 2026-09-03

H200 vs H200 NVL: specs, price per hour, which to rent

NVIDIA H200 (141 GB, $3.058/hr) against NVIDIA H200 NVL (141 GB, $2.567/hr): public specs side by side, live fixed prices from our sheet, what fits in each card's VRAM, and a verdict written for real workloads — both are rentable right now.

H200 vs H200 NVL specifications

Public NVIDIA figures (dense, non-sparsity). The last column is H200 relative to H200 NVL.

SpecH200H200 NVLDifference
ArchitectureHopper (2023)Hopper (2024)
VRAM141 GB HBM3e141 GB HBM3esame
Memory bandwidth4,800 GB/s4,800 GB/ssame
FP16 tensor (dense)990 TFLOPS835 TFLOPS+19%
FP3267.0 TFLOPS60.0 TFLOPS+12%
CUDA cores16,89616,896same
TDP700 W600 W+17%
PCIe · NVLinkGen 5.0 · no NVLinkGen 5.0 · NVLink
PowerScore (RTX 3090 = 100)697588+19%
Max GPUs per machine

H200 vs H200 NVL price per hour

Fixed rates from our sheet — every on-demand price is the marketplace median × 0.70, rounded down. Monthly = 730 hours.

RateH200H200 NVLCheaper
On-demand, per GPU-hour$3.058$2.567H200 NVL (−16%)
Interruptible, per GPU-hour$1.529$1.283H200 NVL
Reserved (3 mo), per GPU-hour$1.987$1.668H200 NVL
On-demand, per month$2,232$1,874H200 NVL
Market median (reference)$4.37$3.67
$ per 1,000 FP16 TFLOP-hours$3.09$3.07H200 NVL (better value)
$ per GB of VRAM per hour$0.0217$0.0182H200 NVL (better value)

Try a full month with storage and bandwidth in the GPU cost calculator.

What fits in VRAM: 141 GB vs 141 GB

WorkloadH200H200 NVL
Largest LLM in FP16, one card~49B~49B
Largest LLM at 4-bit, one card~141B~141B
Flux dev (FP8, ~17 GB)fitsfits
Wan 2.x 14B video (offloaded)yesyes
70B 4-bit LLM on one cardyesyes

Rules of thumb: ~2.4 GB per billion parameters in FP16 all-in, ~0.62 GB in 4-bit. Full tables in the VRAM guide.

Verdict: which should you rent?

Same 141 GB of HBM3e: the SXM H200 offers 700 W, 8-way NVLink and slightly higher clocks; the H200 NVL is a PCIe card in bridged sets of two or four. NVL is cheaper per hour for inference; SXM for large training jobs.

  • Cheaper per hour: H200 NVL ($2.567 vs $3.058, −16%).
  • More VRAM: H200 (141 GB vs 141 GB).
  • More FP16 throughput: H200 (about 1.2×).
  • Best value per TFLOP-hour: H200 NVL.
  • Best value per GB of VRAM: H200 NVL.
  • Multi-GPU: H200 over PCIe · H200 NVL with NVLink.
Is the H200 faster than the H200 NVL?

On dense FP16 tensor throughput the H200 leads by about 1.2× (990 vs 835 TFLOPS). Memory bandwidth matters as much for inference: H200 4,800 GB/s vs H200 NVL 4,800 GB/s.

Which is cheaper to rent, the H200 or the H200 NVL?

The H200 NVL: $2.567/hr on-demand versus $3.058/hr — 16% less. Interruptible rates are $1.529 (H200) and $1.283 (H200 NVL). Per TFLOP-hour the better value is the H200 NVL.

Which has more VRAM and what does that change?

The H200 has 141 GB versus 141 GB. In LLM terms that is roughly a 49B FP16 model (or ~141B in 4-bit) on one card against 49B FP16 (~141B 4-bit). If the model does not fit, speed is irrelevant.

Can I rent both on PowerGPU right now?

Yes — 41 × H200 and 12 × H200 NVL are online as this page renders, deployable in about 30 seconds, billed per second, paid in crypto with no KYC.

Deploy your first GPU in under a minute

Top up in crypto, benchmark us against your current provider. Per-second billing, fixed prices ≥ 30% below market — cancel by just stopping the instance.