The NVIDIA H200 NVL is a data centre GPU designed for generative AI, LLM inference, and high performance computing. With 141 GB of HBM3e and a memory bandwidth of 4,800 GB/s, it processes large models and extensive datasets directly in fast GPU memory. 16,896 CUDA Cores and 528 Tensor Cores provide the computing power for highly parallelised AI and HPC workloads.
- 141 GB HBM3e with ECC for large models and memory-intensive datasets
- 4,800 GB/s memory bandwidth for high data throughput in AI and HPC workloads
- 16,896 CUDA Cores and 528 Tensor Cores for massively parallel computations
- PCIe 5.0 x16 and NVLink for scalable multi-GPU configurations
- Passive dual-slot cooling for use in specially designed GPU servers
141 GB HBM3e for large models and datasets
For large language models and memory-intensive HPC applications, the available GPU memory capacity determines how much data can be processed directly. The H200 NVL provides 141 GB of HBM3e with ECC for this purpose. A memory bandwidth of 4,800 GB/s ensures that the compute units are supplied with data quickly. This is particularly relevant for LLM inference, generative AI, and data-intensive simulations.
Computing power for AI and high performance computing
The Hopper-based GPU combines 16,896 CUDA Cores with 528 Tensor Cores. This provides specialised compute resources for matrix operations and highly parallelised workloads. Supported formats include FP64, FP32, TF32, BF16, FP16, and FP8. This allows you to use the H200 NVL for both scientific calculations and AI models with varying precision requirements.
Designed for scalable GPU servers
The H200 NVL connects to the host system via PCIe 5.0 x16. For multi-GPU systems, it supports configurations with NVIDIA NVLink and up to 900 GB/s connection bandwidth. Up to seven MIG instances, each with 16.5 GB, also enable GPU resources to be allocated to separate workloads. The passive dual-slot cooling requires a server designed for this purpose with sufficient airflow, making it particularly suitable for professional data centre environments.
|
Colour
|
|
|
Primary Colour
|
Bronze
|
|
Dimensions
|
|
|
Length / Depth
|
267 mm
|
|
Width
|
40 mm
|
|
Height
|
111 mm
|
|
Weight
|
1217 g
|
|
Graphics Card
|
|
|
GPU Series
|
NVIDIA Server
|
|
GPU Model
|
NVIDIA H200 NVL
|
|
GPU Series Name
|
Server Hopper
|
|
Graphics Chip
|
GH100
|
|
Manufacturing Process
|
5 nm
|
|
Shader Units
|
16896
|
|
Tensor Cores
|
528
|
|
GPU Memory Type
|
HBM3e
|
|
Slot Type Standard
|
PCIe 5.0
|
|
GPU Memory Size (GB)
|
141
|
|
GPU Memory Interface (bit)
|
6144
|
|
GPU Memory Bandwidth (GB/s)
|
4890
|
|
GPU Memory Clock (MHz)
|
1593
|
|
GPU Memory Clock Effective (Gbps)
|
6.4
|
|
Clock Speeds
|
|
|
Max. GPU Clock (Base)
|
1365 MHz
|
|
Max. GPU Clock (Boost)
|
1785 MHz
|
|
Cooling
|
|
|
Number of GPU Fans
|
0
|
|
Water Cooling
|
|
|
Watercooling
|
No
|
|
Internal Ports
|
|
|
Required Power Connectors
|
16-Pin PCIe 5.0 - 12v-2x6 (1x)
|
|
Expansion Slots
|
|
|
Expansion Slots Required
|
2 Slots
|
|
Compatibility
|
|
|
Compatibility Note
|
Please refer to the manufacturer's compatibility information. You can find the latest overview on the manufacturer's website.
|
|
AI
|
|
|
AI Acceleration
|
Yes
|
|
AI Accelerator
|
Tensor Cores
|
|
Lighting
|
|
|
Lighting / RGB
|
No
|
Product reviews displayed are aggregated from all Pro Gamers Group websites.
| Manufacturer Information | |
| Responsible Person |