NVIDIAVendor documented
H100 NVL
Hopper · Hopper · PCIe (NVL pair) · 2023
Near-SXM Hopper compute in a PCIe part with 94 GB HBM3, tuned for large-model inference; typically deployed as bridged NVL pairs.
LLM inferenceLLM trainingFine-tuningRAG
Precision fingerprint
6432t3216BF168i8
Memory
94 GB
HBM3
Bandwidth
3.9 TB/s
peak
TDP
400 W
air
Max model
~35B
FP16, planning est.
Compute throughput
| FP64 | 67 TFLOPS |
| FP32 | 67 TFLOPS |
| TF32 | 494 TFLOPS |
| FP16 | 989 TFLOPS |
| BF16 | 989 TFLOPS |
| FP8 | 1.98 PFLOPS |
| INT8 | 1.98 PFLOPS |
Platform & software
InterconnectNVLink bridge — 600 GB/s
PCIePCIe 5.0 x16
Coolingair
MIGSupported
PartitioningUp to 7× MIG
VirtualizationvGPU, MIG
FrameworksCUDA, TensorRT-LLM, Triton, NeMo, vLLM
AvailabilityAzure, GCP, OCI, bare-metal
Known limitations
- ·Sold and deployed as bridged NVL pairs
- ·Higher per-card power than H100 PCIe