NVIDIAVendor documented
H200 SXM
Hopper · Hopper (refresh) · SXM5 · 2024
H100 compute with 141 GB HBM3e and 4.8 TB/s. The extra memory notably improves long-context and larger-model inference.
LLM trainingLLM inferenceRAGMultimodalHPC
Precision fingerprint
6432t3216BF168i8
Memory
141 GB
HBM3e
Bandwidth
4.8 TB/s
peak
TDP
700 W
air or liquid
Max model
~55B
FP16, planning est.
Compute throughput
| FP64 | 67 TFLOPS |
| FP32 | 67 TFLOPS |
| TF32 | 494 TFLOPS |
| FP16 | 989 TFLOPS |
| BF16 | 989 TFLOPS |
| FP8 | 1.98 PFLOPS |
| INT8 | 1.98 PFLOPS |
Platform & software
InterconnectNVLink 4 — 900 GB/s
PCIePCIe 5.0 x16
Coolingair or liquid
MIGSupported
PartitioningUp to 7× MIG
VirtualizationvGPU, MIG
FrameworksCUDA, TensorRT-LLM, Triton, NeMo, vLLM
AvailabilityAWS, Azure, GCP, OCI, bare-metal
Known limitations
- ·Same compute as H100 — gains are memory capacity and bandwidth only