GPU comparison · Hopper vs Blackwell
NVIDIA H100 vs B200.
Hopper against Blackwell. Per GPU, the B200 has 2.25× the memory (180 GB), 2.4× the bandwidth (8 TB/s), about 2.3× the dense FP8 and BF16 throughput and FP4 support. The H100 costs about half as much per month and runs every framework out of the box.
- 80 vs 180 GBGPU memory
- 3.35 vs 8 TB/smemory bandwidth
- $1,779 vs $3,439per GPU, per month
H100 vs B200: which one to rent.
Choose the H100 if
- Your models fit in 80 GB per GPU and you want the lowest price per GPU: $1,779 a month.
- You fine-tune or serve small and mid-size models on a proven, fully tuned software stack.
- You need FP64 tensor math (67 TFLOPS on Hopper).
Choose the B200 if
- You need throughput per GPU for training or high-traffic inference.
- Your models need more than 80 GB per GPU, or you want FP4.
- You scale over 8 GPUs: NVLink 5 at 1.8 TB/s per GPU, twice the H100.
H100 vs B200 specs, side by side.
Figures from NVIDIA’s published specifications, per GPU. Tensor figures are peak dense throughput.
| Per GPU | NVIDIA H100 | NVIDIA B200 | B200 / H100 |
|---|---|---|---|
| Architecture | Hopper, TSMC 4N | Blackwell, TSMC 4NP | — |
| Form factor | SXM5 module | SXM module (HGX B200) | — |
| GPU memory | 80 GB HBM3, 5,120-bit | 180 GB HBM3e | 2.3× |
| Memory bandwidth | 3.35 TB/s | Up to 8 TB/s | 2.4× |
| GPU-to-GPU link | NVLink 4, 900 GB/s per GPU | NVLink 5, 1.8 TB/s per GPU | 2× |
| FP4 tensor (dense) | Not supported | 9.0 PFLOPS | — |
| FP8 tensor (dense) | 1,979 TFLOPS | 4,500 TFLOPS | 2.3× |
| FP16 / BF16 tensor (dense) | 989 TFLOPS | 2,250 TFLOPS | 2.3× |
| FP64 | 34 / 67 TFLOPS | 37 TFLOPS | 1.09× |
| Max power | Up to 700 W, configurable | Up to 1,000 W, configurable | — |
| GPUs per server | 1, 2, 4, 8 | 1, 2, 4, 8 | — |
| Per GPU in our servers | 20 vCPU · 200 GB RAM · 2 TB | 28 vCPU · 288 GB RAM · 3.84 TB | — |
Sources: NVIDIA H100 · NVIDIA HGX · NVIDIA DGX B200. Full sheets: H100 · B200
H100 vs B200 rental price.
Our monthly prices, set 30% below the market median of public on-demand prices and paid in crypto. Hourly figures are the monthly price divided by 730 hours.
| CryptGPU | NVIDIA H100 | NVIDIA B200 | B200 / H100 |
|---|---|---|---|
| Price per GPU, per month | $1,779 | $3,439 | 1.93× |
| Equivalent per GPU-hour | $2.44 | $4.71 | — |
| Per GB of GPU memory, per month | $22.24 | $19.11 | 0.86× |
| Per PFLOPS of dense FP8, per month | $899 | $764 | 0.85× |
| Largest server | 8× · $14,232/mo | 8× · $27,512/mo | — |
| Market median, per GPU | $2,548 | $4,920 | — |
| Below the median | −30% | −30% | — |
Same price per GPU at every server size, no hourly metering. How we compare prices · Configure an H100 server · Configure a B200 server
Which models fit on each.
Smallest server that holds each model, by precision. Rule of thumb with headroom for the KV cache, 92% of GPU memory usable.
| Model (total parameters) | H100 · FP8 | B200 · FP8 | H100 · 4-bit | B200 · 4-bit |
|---|---|---|---|---|
| Qwen3.5-9B | 1× | 1× | 1× | 1× |
| Gemma 4 31B | 1× | 1× | 1× | 1× |
| Llama 3.3 70B | 2× | 1× | 1× | 1× |
| gpt-oss-120b | 2× | 1× | 2× | 1× |
| DeepSeek-V4-Flash | 8× | 4× | 4× | 2× |
| Qwen3.5-397B-A17B | 8× | 4× | 4× | 2× |
| DeepSeek-R1 (671B) | — | 8× | 8× | 4× |
| Kimi K2.6 (1T) | — | 8× | — | 4× |
FP8 ≈ 1.2 bytes and 4-bit ≈ 0.65 bytes per parameter. — = larger than the biggest server of that GPU. Size another model · How much VRAM does an LLM need?
H100 vs B200 FAQ.
More on each GPU: NVIDIA H100 · NVIDIA B200.
How much faster is the B200 than the H100?
On paper, about 2.3× the dense FP8 and BF16 throughput and 2.4× the memory bandwidth per GPU. Real speed-ups depend on the model, precision and software; we do not publish benchmarks.
Which is cheaper per unit of compute?
The B200: about $764 per PFLOPS of dense FP8 per month, against $899 on the H100.
Is the H100 still a good buy in 2026?
Yes for fine-tuning and serving models that fit in 80 GB: the software is mature and the price per GPU is low. For new large-scale training, Blackwell gets more done per GPU.
Other GPU comparisons.
Rent the H100 or the B200.
Dedicated servers with 1 to 8 GPUs, one monthly price, paid in crypto. Online in under 10 minutes, no KYC.
