NVIDIA H100 PCIe

Data center GPU · Hopper

Summary

The lowest-power variant of NVIDIA's flagship Hopper-generation data center GPU, with a PCIe interface.

For multi-GPU training with the H100, see the H100 SXM.

For an improved version of this GPU with more VRAM, see the H200 NVL.

Launched
Q4 2022
VRAM
80 GB
Mem. bandwidth
2,000 GB/s
On-demand from
Compare prices from 9 providers →

Tech specs

VRAM 80 GB HBM2e
Memory bandwidth 2,000 GB/s
Interface PCIe
CUDA cores 14,592
Tensor cores 456 (Gen 4)
TDP 350 W
Supported data types
FP64FP32FP16BF16FP8INT8

Cloud rental prices

9 providers · available in 6 countries · lowest price per GPU, per hour

Prices updated

On-demand
$2.50 – $3.39 /GPU/h
Reserved
$1.18 – $3.88 /GPU/h
Compare all NVIDIA H100 PCIe cloud providers & configurations →

Models that fit in VRAM

Total VRAM From 16-bit inference 8-bit inference 4-bit inference
1× H100 PCIe 80 GB $1.18/h GLM-4.7-FlashQwen3-Coder-30B-A3B-InstructNVIDIA-Nemotron-3-Nano-30B-A3B-BF16 GLM-4.7-FlashQwen3-Coder-30B-A3B-InstructNVIDIA-Nemotron-3-Nano-30B-A3B-BF16 gpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16GLM-4.5-Air
2× H100 PCIe 160 GB $3.50/h GLM-4.7-FlashQwen3-Coder-30B-A3B-InstructNVIDIA-Nemotron-3-Nano-30B-A3B-BF16 gpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16GLM-4.5-Air MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b
3× H100 PCIe 240 GB $8.67/h GLM-4.5-AirQwen3-Coder-NextNVIDIA-Nemotron-3-Nano-30B-A3B-BF16 DeepSeek-V4-Flashgpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16 MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b
4× H100 PCIe 320 GB $7.00/h gpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16GLM-4.5-Air MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b
5× H100 PCIe 400 GB $14.45/h DeepSeek-V4-Flashgpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16 MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b DeepSeek-V3.2MiniMax-M2.7gpt-oss-120b
6× H100 PCIe 480 GB $17.34/h DeepSeek-V4-Flashgpt-oss-120bNVIDIA-Nemotron-3-Super-120B-A12B-BF16 MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b DeepSeek-V4-ProMiniMax-M2.7gpt-oss-120b
7× H100 PCIe 560 GB $15.50/h MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b DeepSeek-V4-ProMiniMax-M2.7gpt-oss-120b
8× H100 PCIe 640 GB $14.00/h MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b MiniMax-M2.7DeepSeek-V4-Flashgpt-oss-120b Kimi-K2-Instruct-0905DeepSeek-V4-ProMiniMax-M2.7

Open-weights models that fit in VRAM. Estimated using 🤗 accelerate, plus approximation for up to 8K context.

Peak theoretical performance

FP8 Tensor Core 1,513 TFLOPS
INT8 Tensor Core 1,513 TOPS
BF16 Tensor Core 756 TFLOPS
FP16 Tensor Core 756 TFLOPS
TF32 Tensor Core 378 TFLOPS
FP32 51 TFLOPS
FP64 26 TFLOPS
FP64 Tensor Core 51 TFLOPS

Performance figures assume no sparsity; in cases where only sparse performance figures are published by the manufacturer, these are halved to give approximate dense performance.

Resources

Detailed documentation from the manufacturer.

Similar GPUs

Other accelerators you might compare.