Price floor Every GPU at least 30% below the market median — re-checked against the marketplace weekly.

See the proof

Compare · prices checked 2026-09-03

H100 SXM vs B200: specs, price per hour, which to rent

NVIDIA H100 SXM (80 GB, $1.587/hr) against NVIDIA B200 (192 GB, $4.204/hr): public specs side by side, live fixed prices from our sheet, what fits in each card's VRAM, and a verdict written for real workloads — both are rentable right now.

H100 SXM vs B200 specifications

Public NVIDIA figures (dense, non-sparsity). The last column is H100 SXM relative to B200.

SpecH100 SXMB200Difference
ArchitectureHopper (2022)Blackwell (2024)
VRAM80 GB HBM3192 GB HBM3e−58%
Memory bandwidth3,350 GB/s8,000 GB/s−58%
FP16 tensor (dense)990 TFLOPS2,250 TFLOPS−56%
FP3267.0 TFLOPS
CUDA cores16,896
TDP700 W1000 W−30%
PCIe · NVLinkGen 5.0 · NVLinkGen 5.0 · NVLink
PowerScore (RTX 3090 = 100)6971585−56%
Max GPUs per machine

H100 SXM vs B200 price per hour

Fixed rates from our sheet — every on-demand price is the marketplace median × 0.70, rounded down. Monthly = 730 hours.

RateH100 SXMB200Cheaper
On-demand, per GPU-hour$1.587$4.204H100 SXM (−62%)
Interruptible, per GPU-hour$0.793$2.102H100 SXM
Reserved (3 mo), per GPU-hour$1.031$2.732H100 SXM
On-demand, per month$1,159$3,069H100 SXM
Market median (reference)$2.27$6.01
$ per 1,000 FP16 TFLOP-hours$1.60$1.87H100 SXM (better value)
$ per GB of VRAM per hour$0.0198$0.0219H100 SXM (better value)

Try a full month with storage and bandwidth in the GPU cost calculator.

What fits in VRAM: 80 GB vs 192 GB

WorkloadH100 SXMB200
Largest LLM in FP16, one card~32B~72B
Largest LLM at 4-bit, one card~123B~235B
Flux dev (FP8, ~17 GB)fitsfits
Wan 2.x 14B video (offloaded)yesyes
70B 4-bit LLM on one cardyesyes

Rules of thumb: ~2.4 GB per billion parameters in FP16 all-in, ~0.62 GB in 4-bit. Full tables in the VRAM guide.

Verdict: which should you rent?

Blackwell delivers about 2.3× the dense FP16 throughput of the H100 SXM with 2.4× the memory; the H100 is cheaper per hour and far deeper in supply. B200 for wall-clock time and FP4 serving, H100 SXM for price per FLOP and interruptible availability.

  • Cheaper per hour: H100 SXM ($1.587 vs $4.204, −62%).
  • More VRAM: B200 (192 GB vs 80 GB).
  • More FP16 throughput: B200 (about 2.3×).
  • Best value per TFLOP-hour: H100 SXM.
  • Best value per GB of VRAM: H100 SXM.
  • Multi-GPU: H100 SXM with NVLink · B200 with NVLink.
Is the B200 faster than the H100 SXM?

On dense FP16 tensor throughput the B200 leads by about 2.3× (2,250 vs 990 TFLOPS). Memory bandwidth matters as much for inference: H100 SXM 3,350 GB/s vs B200 8,000 GB/s.

Which is cheaper to rent, the H100 SXM or the B200?

The H100 SXM: $1.587/hr on-demand versus $4.204/hr — 62% less. Interruptible rates are $0.793 (H100 SXM) and $2.102 (B200). Per TFLOP-hour the better value is the H100 SXM.

Which has more VRAM and what does that change?

The B200 has 192 GB versus 80 GB. In LLM terms that is roughly a 72B FP16 model (or ~235B in 4-bit) on one card against 32B FP16 (~123B 4-bit). If the model does not fit, speed is irrelevant.

Can I rent both on PowerGPU right now?

Yes — 69 × H100 SXM and 53 × B200 are online as this page renders, deployable in about 30 seconds, billed per second, paid in crypto with no KYC.

Deploy your first GPU in under a minute

Top up in crypto, benchmark us against your current provider. Per-second billing, fixed prices ≥ 30% below market — cancel by just stopping the instance.