NVIDIA H200 NVL Tensor Core Graphics Card, Supercharging AI and HPC workloads, High-Performance Computing, TDP Up to 600W
Quick Overview
- Llama2 70B Inference 1.9X Faster
- GPT-3 175B Inference 1.6X Faster
- High-Performance Computing 110X Faster
AED 169,999.00 Original price was: AED 169,999.00.AED 149,999.00Current price is: AED 149,999.00.
The GPU for Generative AI and HPC
The NVIDIA H200 GPU supercharges generative AI and high-performance computing (HPC) workloads with game-changing performance and memory capabilities. As the first GPU with HBM3E, the H200’s larger and faster memory fuels the acceleration of generative AI and large language models (LLMs) while advancing scientific computing for HPC workloads.
Higher Performance With Larger, Faster Memory
Based on the NVIDIA Hopper™ architecture, the NVIDIA H200 is the first GPU to offer 141 gigabytes (GB) of HBM3e memory at 4.8 terabytes per second (TB/s) —that’s nearly double the capacity of the NVIDIA H100 GPU with 1.4X more memory bandwidth. The H200’s larger and faster memory accelerates generative AI and LLMs, while advancing scientific computing for HPC workloads with better energy efficiency and lower total cost of ownership.
Unlock Insights With High-Performance LLM Inference
In the ever-evolving landscape of AI, businesses rely on LLMs to address a diverse range of inference needs. An AI inference accelerator must deliver the highest throughput at the lowest TCO when deployed at scale for a massive user base. The H200 boosts inference speed by up to 2X compared to H100 GPUs when handling LLMs like Llama2.
Supercharge High-Performance Computing
Memory bandwidth is crucial for HPC applications as it enables faster data transfer, reducing complex processing bottlenecks. For memory-intensive HPC applications like simulations, scientific research, and artificial intelligence, the H200’s higher memory bandwidth ensures that data can be accessed and manipulated efficiently, leading up to 110X faster time to results compared to CPUs.
Reduce Energy and TCO
With the introduction of the H200, energy efficiency and TCO reach new levels. This cutting-edge technology offers unparalleled performance, all within the same power profile as the H100. AI factories and supercomputing systems that are not only faster but also more eco-friendly, deliver an economic edge that propels the AI and scientific community forward.
Accelerating AI Acceleration for Mainstream Enterprise Servers With H200 NVL

NVIDIA H200 NVL is ideal for lower-power, air-cooled enterprise rack designs that require flexible configurations, delivering acceleration for every AI and HPC workload regardless of size. With up to four GPUs connected by NVIDIA NVLink™ and a 1.5x memory increase, large language model (LLM) inference can be accelerated up to 1.7x, and HPC applications achieve up to 1.3x more performance over the H100 NVL.
Enterprise-Ready: AI Software Streamlines Development and Deployment
NVIDIA H200 NVL comes with a five-year NVIDIA Enterprise subscription. This subscription includes NVIDIA AI Enterprise to simplify the way you build an enterprise AI-ready platform. H200 accelerates AI development and deployment for production-ready generative AI solutions, including computer vision, speech AI, retrieval augmented generation (RAG), and more. NVIDIA AI Enterprise includes NVIDIA NIM™, a set of easy-to-use microservices designed to speed up enterprise generative AI deployment. Together, deployments have enterprise-grade security, manageability, stability, and support. This results in performance-optimized AI solutions that deliver faster business value and actionable insights.

NVIDIA H200 NVL Specifications
- FP64 Performance: 30 TFLOPS
- FP64 Tensor Core Performance: 60 TFLOPS
- FP32 Performance: 60 TFLOPS
- TF32 Tensor Core Performance: 835 TFLOPS²
- BFLOAT16 Tensor Core Performance: 1,671 TFLOPS²
- FP16 Tensor Core Performance: 1,671 TFLOPS²
- FP8 Tensor Core Performance: 3,341 TFLOPS²
- INT8 Tensor Core Performance: 3,341 TFLOPS²
- GPU Memory: 141GB
- GPU Memory Bandwidth: 4.8TB/s
- Decoders:
- 7 NVDEC
- 7 JPEG
- Confidential Computing: Supported
- Maximum Thermal Design Power (TDP): Up to 600W (configurable)
- Multi-Instance GPU (MIG) Support: Up to 7 MIGs @ 16.5GB each
- Form Factor:
- PCIe
- Dual-slot air-cooled
- Interconnect:
- 2- or 4-way NVIDIA NVLink bridge: 900GB/s per GPU
- PCIe Gen5: 128GB/s
- Server Options:
- NVIDIA MGX™ H200 NVL partner systems
- NVIDIA-Certified Systems™ with up to 8 GPUs
- NVIDIA AI Enterprise: Included
| Standard Delivery | Two Days Delivery |
|---|---|
| Included Warranty | 1 Year Free Warranty |
Product Reviews
You must be logged in to post a review.
Product Additional information
| Standard Delivery | Two Days Delivery |
|---|---|
| Included Warranty | 1 Year Free Warranty |
Related products
Cooler Master ML 240 Atmos II LCD ARGB Simple Water Cooling CPU Cooler, White, MLX-D24M-A25SZ-LW
In stock
ASUS ProArt Z890-CREATOR WIFI Intel Z890 LGA1851 ATX content creation motherboard with PCIe® 5.0, DDR5 slots, two USB4® ports, 10G and 2.5G Ethernet, WiFi 7, five M.2 slots, plus a USB 20Gpbs front-panel header
In stock
ASUS ROG Strix Helios II EATX mid-tower gaming case with dual tempered glass side panels, GPU support for up to 450mm in length, aluminum frame and front panel, GPU braces and 420mm radiator support, Black
In stock
MSI PRO B760M-E Motherboard, Lightning Gen 4 PCI-e with Steel Armor, Supports DDR5 Memory, Dual Channel DDR5 6400+MHz (OC), Black
In stock
ASUS PRIME Z790-P WIFI Motherboard, PCIe Gen5 Slot, USB 3.2 Gen 2×2 Type-C® & Thunderbolt (USB4®) Support, WiFi 6, Al Cooling II
In stock
ASUS PRIME B860M-K Motherboard, AEMP III, USB 10Gbps Type-A, USB 5Gbps Type-C, PCIe 5.0 M.2, 2.5Gb Ethernet, Aura Sync
In stock
ASUS Pro WS WRX90E-SAGE SE AMD sTR5 EEB workstation motherboard, 7 PCIe 5.0 x16 slots, multi-GPU support, robust 32+3+3+3 power-stage design, CPU and memory overclocking ready, 4 PCIe 5.0 M.2 slots, dual 10 Gb LAN, PCIe Q-Release Slim
In stock

Reviews
Clear filtersThere are no reviews yet.