NVIDIA H200 NVL

Data center GPU · Hopper

Summary

A lower-power variant of NVIDIA's H200 GPU, the NVIDIA H200 NVL improves on the H100 NVL with a 1.5x memory increase and 1.2x bandwidth increase, leading to up to 1.7x faster inference.

Launched
Q4 2024
VRAM
141 GB
Mem. bandwidth
4,800 GB/s
On-demand from
Compare prices from 3 providers →

Tech specs

VRAM 141 GB HBM3e
Memory bandwidth 4,800 GB/s
Interface PCIe
CUDA cores 14,592
Tensor cores 456 (Gen 4)
TDP 600 W
Supported data types
FP64FP32FP16BF16FP8INT8

Cloud rental prices

3 providers · lowest price per GPU, per hour

Prices updated

On-demand
$3.44 – $3.82 /GPU/h
Reserved
$3.43 – $4.54 /GPU/h
Provider On-demand Reserved Spot
Cirrascale $3.43
DigitalOcean $3.44
Massed Compute $3.62
Compare all NVIDIA H200 NVL cloud providers & configurations →

Models that fit in VRAM

Open-weights models that fit in VRAM. Estimated using 🤗 accelerate, plus approximation for up to 8K context.

Media

Peak theoretical performance

FP8 Tensor Core 1,670 TFLOPS
INT8 Tensor Core 1,670 TOPS
BF16 Tensor Core 835 TFLOPS
FP16 Tensor Core 835 TFLOPS
TF32 Tensor Core 417 TFLOPS
FP32 60 TFLOPS
FP64 30 TFLOPS
FP64 Tensor Core 60 TFLOPS

Performance figures assume no sparsity; in cases where only sparse performance figures are published by the manufacturer, these are halved to give approximate dense performance.

Resources

Detailed documentation from the manufacturer.

Similar GPUs

Other accelerators you might compare.