Price floor Every GPU at least 30% below the market median — re-checked against the marketplace weekly.

See the proof

Compare · prices checked 2026-09-14

H200 vs H200 NVL: specs, price per hour, which to rent

NVIDIA H200 (141 GB, $2.791/hr) against NVIDIA H200 NVL (141 GB, $2.650/hr): public specs side by side, live fixed prices from our sheet, what fits in each card's VRAM, and a verdict written for real workloads — both are rentable right now.

H200 vs H200 NVL specifications

Public NVIDIA figures (dense, non-sparsity). The last column is H200 relative to H200 NVL.

SpecH200H200 NVLDifference
ArchitectureHopper (2023)Hopper (2024)
VRAM141 GB HBM3e141 GB HBM3esame
Memory bandwidth4,800 GB/s4,800 GB/ssame
FP16 tensor (dense)990 TFLOPS835 TFLOPS+19%
FP3267.0 TFLOPS60.0 TFLOPS+12%
CUDA cores16,89616,896same
TDP700 W600 W+17%
PCIe · NVLinkGen 5.0 · no NVLinkGen 5.0 · NVLink
PowerScore (RTX 3090 = 100)697588+19%
Max GPUs per machine

H200 vs H200 NVL price per hour

Fixed rates from our sheet — every on-demand price is the marketplace median × 0.70, rounded down. Monthly = 730 hours.

RateH200H200 NVLCheaper
On-demand, per GPU-hour$2.791$2.650H200 NVL (−5%)
Interruptible, per GPU-hour$1.395$1.325H200 NVL
Reserved (3 mo), per GPU-hour$1.814$1.722H200 NVL
On-demand, per month$2,037$1,935H200 NVL
Market median (reference)$3.99$3.79
$ per 1,000 FP16 TFLOP-hours$2.82$3.17H200 (better value)
$ per GB of VRAM per hour$0.0198$0.0188H200 NVL (better value)

Try a full month with storage and bandwidth in the GPU cost calculator.

What fits in VRAM: 141 GB vs 141 GB

WorkloadH200H200 NVL
Largest LLM in FP16, one card~49B~49B
Largest LLM at 4-bit, one card~141B~141B
Flux dev (FP8, ~17 GB)fitsfits
Wan 2.x 14B video (offloaded)yesyes
70B 4-bit LLM on one cardyesyes

Rules of thumb: ~2.4 GB per billion parameters in FP16 all-in, ~0.62 GB in 4-bit. Full tables in the VRAM guide.

Verdict: which should you rent?

Same 141 GB of HBM3e: the SXM H200 offers 700 W, 8-way NVLink and slightly higher clocks; the H200 NVL is a PCIe card in bridged sets of two or four. NVL is cheaper per hour for inference; SXM for large training jobs.

  • Cheaper per hour: H200 NVL ($2.650 vs $2.791, −5%).
  • More VRAM: H200 (141 GB vs 141 GB).
  • More FP16 throughput: H200 (about 1.2×).
  • Best value per TFLOP-hour: H200.
  • Best value per GB of VRAM: H200 NVL.
  • Multi-GPU: H200 over PCIe · H200 NVL with NVLink.
Is the H200 faster than the H200 NVL?

On dense FP16 tensor throughput the H200 leads by about 1.2× (990 vs 835 TFLOPS). Memory bandwidth matters as much for inference: H200 4,800 GB/s vs H200 NVL 4,800 GB/s.

Which is cheaper to rent, the H200 or the H200 NVL?

The H200 NVL: $2.650/hr on-demand versus $2.791/hr — 5% less. Interruptible rates are $1.395 (H200) and $1.325 (H200 NVL). Per TFLOP-hour the better value is the H200.

Which has more VRAM and what does that change?

The H200 has 141 GB versus 141 GB. In LLM terms that is roughly a 49B FP16 model (or ~141B in 4-bit) on one card against 49B FP16 (~141B 4-bit). If the model does not fit, speed is irrelevant.

Can I rent both on PowerGPU right now?

Yes — 50 × H200 and 27 × H200 NVL are online as this page renders, deployable in about 30 seconds, billed per second, paid in crypto with no KYC.

Deploy your first GPU in under a minute

Top up in crypto, benchmark us against your current provider. Per-second billing, fixed prices ≥ 30% below market — cancel by just stopping the instance.