NVIDIA H100 SXM

Data center GPU · Hopper

Summary

The most powerful variant of NVIDIA's flagship Hopper-generation data center GPU – ideal for multi-GPU training runs.

For an improved version of this GPU with more VRAM, see the H200 SXM.

For newer Blackwell-generation GPUs, see the B200 or B300.

Launched
Q4 2022
VRAM
80 GB
Mem. bandwidth
3,350 GB/s
On-demand from
Compare prices from 23 providers →

Tech specs

VRAM 80 GB HBM3
Memory bandwidth 3,350 GB/s
Interface SXM
CUDA cores 16,896
Tensor cores 528 (Gen 4)
TDP 700 W
Supported data types
FP64FP32FP16BF16FP8INT8

Cloud rental prices

23 providers · available in 29 countries · lowest price per GPU, per hour

Models that fit in VRAM

Total VRAM From 16-bit inference 8-bit inference 4-bit inference
1× H100 SXM 80 GB $1.84/h GLM-4.7-FlashQwen3-Coder-30B-A3B-InstructNVIDIA-Nemotron-3-Nano-30B-A3B-BF16 GLM-4.7-FlashQwen3-Coder-30B-A3B-InstructNVIDIA-Nemotron-3-Nano-30B-A3B-BF16 gpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16GLM-4.5-Air
2× H100 SXM 160 GB $3.68/h GLM-4.7-FlashQwen3-Coder-30B-A3B-InstructNVIDIA-Nemotron-3-Nano-30B-A3B-BF16 gpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16GLM-4.5-Air MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b
3× H100 SXM 240 GB $8.97/h GLM-4.5-AirQwen3-Coder-NextNVIDIA-Nemotron-3-Nano-30B-A3B-BF16 DeepSeek-V4-Flashgpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16 MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b
4× H100 SXM 320 GB $7.36/h gpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16GLM-4.5-Air MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b
5× H100 SXM 400 GB $14.95/h DeepSeek-V4-Flashgpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16 MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b DeepSeek-V3.2MiniMax-M2.7gpt-oss-120b
6× H100 SXM 480 GB $17.94/h DeepSeek-V4-Flashgpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16 MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b DeepSeek-V4-ProMiniMax-M2.7gpt-oss-120b
7× H100 SXM 560 GB $20.93/h MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b DeepSeek-V4-ProMiniMax-M2.7gpt-oss-120b
8× H100 SXM 640 GB $14.73/h MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b Kimi-K2-Instruct-0905DeepSeek-V4-ProMiniMax-M2.7

Open-weights models that fit in VRAM. Estimated using 🤗 accelerate, plus approximation for up to 8K context.

Media

Peak theoretical performance

FP8 Tensor Core 1,979 TFLOPS
INT8 Tensor Core 1,979 TOPS
BF16 Tensor Core 990 TFLOPS
FP16 Tensor Core 990 TFLOPS
TF32 Tensor Core 494 TFLOPS
FP32 67 TFLOPS
FP64 34 TFLOPS
FP64 Tensor Core 67 TFLOPS

Performance figures assume no sparsity; in cases where only sparse performance figures are published by the manufacturer, these are halved to give approximate dense performance.

Resources

Detailed documentation from the manufacturer.

Similar GPUs

Other accelerators you might compare.