H100

Values shown for H100 SXM (the variant deployed in Ferranti's h100-ferranti partitions). * denotes "with sparsity" — without sparsity values are half.

Spec H100 SXM H100 PCIe
Architecture NVIDIA Hopper (GH100) NVIDIA Hopper (GH100)
Process TSMC 4N TSMC 4N
Transistors 80 B 80 B
FP64 34 TFLOPS 26 TFLOPS
FP64 Tensor Core 67 TFLOPS 51 TFLOPS
FP32 67 TFLOPS 51 TFLOPS
TF32 Tensor Core 989* TFLOPS 756* TFLOPS
BFLOAT16 Tensor Core 1 979* TFLOPS 1 513* TFLOPS
FP16 Tensor Core 1 979* TFLOPS 1 513* TFLOPS
FP8 Tensor Core 3 958* TFLOPS 3 026* TFLOPS
INT8 Tensor Core 3 958* TOPS 3 026* TOPS
GPU memory 80 GB HBM3 80 GB HBM2e
Memory bandwidth 3.35 TB/s 2 TB/s
Decoders 7 NVDEC, 7 JPEG 7 NVDEC, 7 JPEG
Max TDP up to 700 W (configurable) 300–350 W (configurable)
Multi-Instance GPU up to 7 MIGs @ 10 GB up to 7 MIGs @ 10 GB
Form factor SXM PCIe (dual-slot, air-cooled)
Interconnect NVLink 900 GB/s, PCIe Gen5 128 GB/s NVLink 600 GB/s, PCIe Gen5 128 GB/s

Architecture detail (from NVIDIA Hopper whitepaper / official product page): 528 fourth-generation Tensor cores, 16 896 FP32 / 8 448 FP64 CUDA cores, with a Transformer Engine that mixes FP8 and FP16 precisions.


Last update: July 17, 2026
Created: July 17, 2026