
Nvidia H200 NVL Graphic Card 141 GB Passive PCIe – 900-21010-0040-000
✔ FP64: 34 TFLOPS
✔ FP64 Tensor Core: 67 TFLOPS
✔ FP32: 67 TFLOPS
✔ TF32 Tensor Core²: 989 TFLOPS
✔ Architecture: Blackwell
✔ BFLOAT16 Tensor Core²: 1,979 TFLOPS
✔ FP16 Tensor Core²: 1,979 TFLOPS
✔ FP8 Tensor Core²: 3,958 TFLOPS
✔ INT8 Tensor Core²: 3,958 TFLOPS
✔ GPU Memory: 141GB
✔ GPU Memory Bandwidth: 4.8TB/s
✔ Decoders: 7 NVDEC, 7 JPEG
✔ Confidential Computing: Supported
✔ Max Thermal Design Power (TDP): Up to 600W (configurable)
✔ Multi-Instance GPUs: Up to 7 MIGs @16.5GB each
✔ Form Factor: PCIe
✔ Interconnect: 2- or 4-way NVIDIA NVLink bridge: 900GB/s, PCIe Gen5: 128GB/s
✔ Server Options: NVIDIA MGX™ H200 NVL partner and NVIDIA-Certified Systems with up to 8 GPUs
✔ NVIDIA AI Enterprise: Add-on
✔ Warranty: 3 years manufacturer parts or replace
Availability subject to compliance verification.
Ships in 2 weeks from payment. All sales final. No returns or cancellations. For bulk inquiries, consult a live chat agent or call our toll-free number.
NVIDIA H200 Tensor Core GPU - PCIe Powerhouse
Breakthrough Performance for AI and Data Center Applications
The NVIDIA H200 Tensor Core GPU in its PCIe form factor offers groundbreaking performance for AI workloads, featuring 141GB of memory and a staggering 4.8TB/s bandwidth. This configuration is optimized for large-scale deployments, supporting up to 8 GPUs per server and utilizing NVLink bridges for high-speed data transfer at 900GB/s. With advanced tensor cores delivering nearly 4,000 TFLOPS in FP8 and INT8 operations, the H200 PCIe is designed for demanding data center environments, scalable AI, and multi-tenant workloads through MIG partitioning for maximum efficiency.

AI Acceleration for Mainstream Enterprise Servers With H200 NVL
NVIDIA H200 NVL is ideal for lower-power, air-cooled enterprise rack designs that require flexible configurations, delivering acceleration for every AI and HPC workload regardless of size. With up to four GPUs connected by NVIDIA NVLink™ and a 1.5X memory increase, large language model (LLM) inference can be accelerated up to 1.7X and HPC applications achieve up to 1.3X more performance over the H100 NVL.

Specification | H200 NVL (PCIe) |
|---|---|
FP64 | 34 TFLOPS |
FP64 Tensor Core | 67 TFLOPS |
FP32 | 67 TFLOPS |
TF32 Tensor Core² | 989 TFLOPS |
BFLOAT16 Tensor Core² | 1,979 TFLOPS |
FP16 Tensor Core² | 1,979 TFLOPS |
FP8 Tensor Core² | 3,958 TFLOPS |
INT8 Tensor Core² | 3,958 TFLOPS |
GPU Memory | 141GB |
GPU Memory Bandwidth | 4.8TB/s |
Decoders | 7 NVDEC, 7 JPEG |
Confidential Computing | Supported |
Max Thermal Design Power (TDP) | Up to 600W (configurable) |
Multi-Instance GPUs | Up to 7 MIGs @16.5GB each |
Form Factor | PCIe |
Interconnect | 2- or 4-way NVIDIA NVLink bridge: 900GB/s, PCIe Gen5: 128GB/s |
Server Options | NVIDIA MGX™ H200 NVL partner and NVIDIA-Certified Systems with up to 8 GPUs |
NVIDIA AI Enterprise | Add-on |