Price floor Every GPU at least 30% below the market median — re-checked against the marketplace weekly.

See the proof

Compare · prices checked 2026-09-03

H200 vs B200: specs, price per hour, which to rent

NVIDIA H200 (141 GB, $3.058/hr) against NVIDIA B200 (192 GB, $4.204/hr): public specs side by side, live fixed prices from our sheet, what fits in each card's VRAM, and a verdict written for real workloads — both are rentable right now.

H200 vs B200 specifications

Public NVIDIA figures (dense, non-sparsity). The last column is H200 relative to B200.

SpecH200B200Difference
ArchitectureHopper (2023)Blackwell (2024)
VRAM141 GB HBM3e192 GB HBM3e−27%
Memory bandwidth4,800 GB/s8,000 GB/s−40%
FP16 tensor (dense)990 TFLOPS2,250 TFLOPS−56%
FP3267.0 TFLOPS
CUDA cores16,896
TDP700 W1000 W−30%
PCIe · NVLinkGen 5.0 · no NVLinkGen 5.0 · NVLink
PowerScore (RTX 3090 = 100)6971585−56%
Max GPUs per machine

H200 vs B200 price per hour

Fixed rates from our sheet — every on-demand price is the marketplace median × 0.70, rounded down. Monthly = 730 hours.

RateH200B200Cheaper
On-demand, per GPU-hour$3.058$4.204H200 (−27%)
Interruptible, per GPU-hour$1.529$2.102H200
Reserved (3 mo), per GPU-hour$1.987$2.732H200
On-demand, per month$2,232$3,069H200
Market median (reference)$4.37$6.01
$ per 1,000 FP16 TFLOP-hours$3.09$1.87B200 (better value)
$ per GB of VRAM per hour$0.0217$0.0219H200 (better value)

Try a full month with storage and bandwidth in the GPU cost calculator.

What fits in VRAM: 141 GB vs 192 GB

WorkloadH200B200
Largest LLM in FP16, one card~49B~72B
Largest LLM at 4-bit, one card~141B~235B
Flux dev (FP8, ~17 GB)fitsfits
Wan 2.x 14B video (offloaded)yesyes
70B 4-bit LLM on one cardyesyes

Rules of thumb: ~2.4 GB per billion parameters in FP16 all-in, ~0.62 GB in 4-bit. Full tables in the VRAM guide.

Verdict: which should you rent?

The B200 roughly doubles FP16 throughput and bandwidth over the H200 and adds FP4; the H200 stays cheaper per hour with 141 GB. For throughput-critical training and FP4 inference the B200 usually costs less per token; for memory capacity at a lower rate, the H200.

  • Cheaper per hour: H200 ($3.058 vs $4.204, −27%).
  • More VRAM: B200 (192 GB vs 141 GB).
  • More FP16 throughput: B200 (about 2.3×).
  • Best value per TFLOP-hour: B200.
  • Best value per GB of VRAM: H200.
  • Multi-GPU: H200 over PCIe · B200 with NVLink.
Is the B200 faster than the H200?

On dense FP16 tensor throughput the B200 leads by about 2.3× (2,250 vs 990 TFLOPS). Memory bandwidth matters as much for inference: H200 4,800 GB/s vs B200 8,000 GB/s.

Which is cheaper to rent, the H200 or the B200?

The H200: $3.058/hr on-demand versus $4.204/hr — 27% less. Interruptible rates are $1.529 (H200) and $2.102 (B200). Per TFLOP-hour the better value is the B200.

Which has more VRAM and what does that change?

The B200 has 192 GB versus 141 GB. In LLM terms that is roughly a 72B FP16 model (or ~235B in 4-bit) on one card against 49B FP16 (~141B 4-bit). If the model does not fit, speed is irrelevant.

Can I rent both on PowerGPU right now?

Yes — 41 × H200 and 53 × B200 are online as this page renders, deployable in about 30 seconds, billed per second, paid in crypto with no KYC.

Deploy your first GPU in under a minute

Top up in crypto, benchmark us against your current provider. Per-second billing, fixed prices ≥ 30% below market — cancel by just stopping the instance.