H200-Ready Models
These popular AI models can run on the H200. Actual requirements vary by model, engine, precision, and configuration.
text
GLM 5.2
753B MoE model with 1M-token context for agentic reasoning, coding, and tool use
textvision
Gemma 4 31B IT
Gemma 4 31B dense vision-language model by Google with 256K context and thinking mode
text
NVIDIA Nemotron 3.5 Lightning 30B A3B BF16
Hybrid Mamba-attention MoE with 3B active params, reasoning control and tool calling
H200 GPU Details
| Category | H200 |
|---|---|
| GPU Name | GH100 |
| Architecture | Hopper |
| Process Size | 5 nm |
| Transistors | 80,000 million |
| Release Date | Nov 18th, 2024 |
| Base Clock | 1365 MHz |
| Boost Clock | 1785 MHz |
| Memory Size | 141 GB |
| Memory Type | HBM3e |
| Bandwidth | 3.36 TB/s |
| Tensor Cores | 528 |
| FP16 (half) | 241.3 TFLOPS (4:1) |
| FP32 (float) | 60.32 TFLOPS |
