NVIDIAEstimated
B200
Blackwell · Blackwell · SXM · 2024
Blackwell generation with second-gen Transformer Engine and FP4. A large step in FP8 throughput and memory over Hopper.
LLM trainingLLM inferenceMultimodalHPC
Precision fingerprint
6432t3216BF168i8
Memory
192 GB
HBM3e
Bandwidth
8 TB/s
peak
TDP
1000 W
liquid
Max model
~80B
FP16, planning est.
Compute throughput
| FP64 | 40 TFLOPS |
| FP32 | 40 TFLOPS |
| TF32 | 1.10 PFLOPS |
| FP16 | 2.25 PFLOPS |
| BF16 | 2.25 PFLOPS |
| FP8 | 4.50 PFLOPS |
| INT8 | 4.50 PFLOPS |
Platform & software
InterconnectNVLink 5 — 1.8 TB/s
PCIePCIe 6.0 x16
Coolingliquid
MIGSupported
PartitioningMIG
VirtualizationMIG, vGPU
FrameworksCUDA, TensorRT-LLM, Triton, NeMo, vLLM
AvailabilityAzure, GCP, OCI, bare-metal
Known limitations
- ·1000 W typically requires direct-liquid cooling
- ·Early-life availability constrained