Space Computing Energy Tech Transport Science Dev Loyalty ↗ ◐ Dark mode ◎ Enable Alerts
Compare · GPUs & AI chips

H200 SXM vs Gaudi 2

RANKED BY DENSE FP16/BF16 TENSOR TFLOPS PER CHIP #8 VS #14 CHECKED 06 SEPT 2026
Side by side
Verdict

H200 SXM ranks #8 of 20 on Dense FP16/BF16 tensor TFLOPS per chip (0.990 PFLOPS); Gaudi 2 ranks #14 (432 TFLOPS).

Rank 08 of 20 · leads
Dense FP16/BF16 tensor TFLOPS per chip
0.990 PFLOPS
MakerNVIDIA
Date2024
Verified04 Sept 2026
Evidence
NVIDIA H200 product page, H200 SXM column: FP16/BF16 Tensor Core 1,979 TFLOPS. Footnote 2: "With sparsity." Dense is 1,979 / 2 = 989.5 TFLOPS (0.990 PFLOPS). 141 GB HBM3e. Same Hopper Tensor Core throughput as H100 SXM; H200's difference is memory.
Rank 14 of 20
Dense FP16/BF16 tensor TFLOPS per chip
432 TFLOPS
MakerIntel
Date2022
Verified04 Sept 2026
Evidence
Intel Gaudi 3 white paper (document 817486) comparison table lists Gaudi 2 "BF16 MME TFLOPS 432" and "FP8 MME TFLOPS 865". This row uses the BF16 MME cell. 96 GB HBM2e. The same comparison table is the official Intel source already used for the Gaudi 3 row.
Try another pair
Source Vendor product pages and official spec documents (AMD Instinct MI355X / MI350X / MI325X / MI300X / MI300A / MI250X / MI250 / MI210; NVIDIA… Last checked 06 Sept 2026
Full GPUs & AI chips ranking Spot an error? Send a tip