Comparison · Blackwell against Hopper, at the top

# B200 SXM6 or H200 SXM5?

Both are dedicated, both are monthly, both open without an identity check. What separates them is memory, bandwidth and $3,385 a month — and which of those matters depends entirely on what you intend to run.

- $11,790/mo B200 SXM6, one card
- $8,405/mo H200 SXM5, one card
- 180 vs 141 GB memory per card
- 1.7× bandwidth gap

## The short answer

Which one, and when

- **Pick the B200 SXM6:** when your model needs more than 141 GB on a single card. That is a threshold, not a preference — below it, the extra $3,385/month buys you nothing.
- **Pick the H200 SXM5:** when 141 GB holds your model. It saves $3,385/month, and memory you do not use is memory you paid for.
- **Neither, if:** your model does not fit on a single card of either. Sharding costs synchronisation at every layer — see the [interconnect guide](https://gpuserver.io/guides/nvlink-vs-pcie) before buying a multi-GPU node.

## Side by side

**B200 SXM6 compared with H200 SXM5**

|  | [B200 SXM6](https://gpuserver.io/gpu/b200) | [H200 SXM5](https://gpuserver.io/gpu/h200) | Difference |
|---|---|---|---|
| Architecture | Blackwell | Hopper | Blackwell is newer |
| Memory per card | 180 GB HBM3e | 141 GB HBM3e | 39 GB in favour of the B200 SXM6 |
| Memory bandwidth Predicts generation speed | 8,000 GB/s | 4,800 GB/s | 1.67× in favour of the B200 SXM6 |
| Form factor | SXM on an HGX board | SXM on an HGX board | Same |
| Node sizes | 4×, 8× | 4×, 8× | Up to 1440 GB in one node |
| From, per month | $11,790 | $8,405 | $3,385/month apart |

## Price at every node size

**Monthly price of both cards at each available node size**

| Node | B200 SXM6 | H200 SXM5 | Total VRAM |  |
|---|---|---|---|---|
| 4 × GPU | $11,790/mo | $8,405/mo | 720 GB / 564 GB | [B200 SXM6](https://gpuserver.io/configure?c=b200-x4) [H200 SXM5](https://gpuserver.io/configure?c=h200-x4) |
| 8 × GPU | $22,351/mo | $15,945/mo | 1440 GB / 1128 GB | [B200 SXM6](https://gpuserver.io/configure?c=b200-x8) [H200 SXM5](https://gpuserver.io/configure?c=h200-x8) |

## What each one holds on a single card

At 8k context, best precision that fits. This is the difference that decides a purchase — not the specification sheet.

### Only the B200 SXM6

Nothing. Every model the B200 SXM6 holds on one card, the H200 SXM5 holds too — the difference between them is speed, not capability.

### Only the H200 SXM5

Nothing. The B200 SXM6 holds everything the H200 SXM5 does.

**The same model on both: Llama 3.3 70B.** On the B200 SXM6, roughly 26 tokens/second at BF16 / FP16; on the H200 SXM5, roughly 31. Single-stream and deliberately conservative — with continuous batching the aggregate is several times higher on both. Treat the *ratio* as the useful number, not the absolute.

## Cost per GB and per TB/s

Two ratios that cut through the specification sheet. The first tells you what memory costs; the second what speed costs.

**Cost per gigabyte of VRAM and per terabyte per second of bandwidth**

| Card | $/GB of VRAM, per month | $/TB per second, per month | Better at |
|---|---|---|---|
| [B200 SXM6](https://gpuserver.io/gpu/b200) | $65.50 | $1,474 | Bandwidth |
| [H200 SXM5](https://gpuserver.io/gpu/h200) | $59.61 | $1,751 | Memory |

These ratios rank cards; they do not choose one. A card that is cheaper per gigabyte is worthless if it has too few gigabytes to hold your model at all — capacity is a threshold, not a slope. Use them to break a tie between two cards that both fit.

## B200 SXM6 against H200 SXM5, in short

Is the B200 SXM6 worth the extra money over the H200 SXM5?

It costs $3,385 more a month — 40% — for 39 GB more memory and 67% more memory bandwidth. Both hold the same set of models on a single card, so the difference is speed rather than capability.

Which one is faster for LLM inference?

Generation speed is bounded by memory bandwidth, because each token requires re-reading the active weights. The B200 SXM6 has 8,000 GB/s against 4,800 GB/s, so the B200 SXM6 leads by roughly 67% on a model both of them hold.

Do both come as multi-GPU nodes?

B200 SXM6: 4×, 8×. H200 SXM5: 4×, 8×. SXM cards live on an HGX board of four or eight — never two, never ten — and are joined by NVLink rather than PCIe.

Can I rent either without an identity check?

Yes, both, on the same terms. No document, no selfie, no phone number and no company registration, at any level of spend. Payment is in cryptocurrency, each invoice gets its own address, and root access arrives under 5 minutes after the first confirmation.

## Other comparisons

- [H200 SXM5 vs H100 SXM5  More memory, more bandwidth, same architecture Compare](https://gpuserver.io/compare/h200-vs-h100-sxm5)
- [H100 SXM5 vs B200 SXM6  Is Blackwell worth the step up? Compare](https://gpuserver.io/compare/h100-sxm5-vs-b200)

## Both are racked and waiting, in all 6 data centres.

Whichever you pick, it is monthly, paid in crypto, opened without an identity check and delivered in under 5 minutes.

[Open the configurator](https://gpuserver.io/configure) [Read the guides](https://gpuserver.io/guides)

---

Source: https://gpuserver.io/compare/b200-vs-h200/. This file is generated from the same data as the website; if a figure here differs from a page, the page is authoritative and this file is stale — the canonical source is https://gpuserver.io/.
