Space Computing Energy Tech Transport Science Dev Loyalty ↗ ◐ Dark mode ◎ Enable Alerts
Compare · GPUs & AI chips

Blackwell Ultra B300 SXM vs Gaudi 2

RANKED BY DENSE FP16/BF16 TENSOR TFLOPS PER CHIP #4 VS #14 CHECKED 06 SEPT 2026
Side by side
Verdict

Blackwell Ultra B300 SXM ranks #4 of 20 on Dense FP16/BF16 tensor TFLOPS per chip (2.25 PFLOPS); Gaudi 2 ranks #14 (432 TFLOPS).

Rank 04 of 20 · leads
Dense FP16/BF16 tensor TFLOPS per chip
2.25 PFLOPS
MakerNVIDIA
Date2025
Verified04 Sept 2026
Evidence
NVIDIA HGX Platform page, HGX B300 row ("8x NVIDIA Blackwell Ultra SXM", shipping now). FP16/BF16 Tensor Core is the same 36 PFLOPS sparse / 18 PFLOPS dense board total as B200, so 2.25 PFLOPS dense per chip. B300's advertised uplift versus B200 is at FP4 and HBM capacity, not at dense FP16/BF16 — so on this metric B200 and B300 are tied.
Rank 14 of 20
Dense FP16/BF16 tensor TFLOPS per chip
432 TFLOPS
MakerIntel
Date2022
Verified04 Sept 2026
Evidence
Intel Gaudi 3 white paper (document 817486) comparison table lists Gaudi 2 "BF16 MME TFLOPS 432" and "FP8 MME TFLOPS 865". This row uses the BF16 MME cell. 96 GB HBM2e. The same comparison table is the official Intel source already used for the Gaudi 3 row.
Try another pair
Source Vendor product pages and official spec documents (AMD Instinct MI355X / MI350X / MI325X / MI300X / MI300A / MI250X / MI250 / MI210; NVIDIA… Last checked 06 Sept 2026
Full GPUs & AI chips ranking Spot an error? Send a tip