Space Computing Energy Tech Transport Science Dev Loyalty ↗ ◐ Dark mode ◎ Enable Alerts
Compare · GPUs & AI chips

Instinct MI300A vs Trainium2

RANKED BY DENSE FP16/BF16 TENSOR TFLOPS PER CHIP #10 VS #12 CHECKED 06 SEPT 2026
Side by side
Verdict

Instinct MI300A ranks #10 of 20 on Dense FP16/BF16 tensor TFLOPS per chip (980.6 TFLOPS); Trainium2 ranks #12 (667 TFLOPS).

Rank 10 of 20 · leads
Dense FP16/BF16 tensor TFLOPS per chip
980.6 TFLOPS
MakerAMD
Date2023
Verified04 Sept 2026
Evidence
AMD Instinct MI300A product page: "Peak Half Precision (FP16) Performance 980.6 TFLOPs". Structured-sparsity counterpart on the same page is 1.96 PFLOPs and was not used. Launch date on the page: 6 December 2023. 128 GB unified HBM3. Dense FP16 sits just below H100/H200 SXM (989.5 TFLOPS).
Rank 12 of 20
Dense FP16/BF16 tensor TFLOPS per chip
667 TFLOPS
MakerAWS
Date2024
Verified04 Sept 2026
Evidence
AWS Neuron Trainium2 Architecture page, compute table: each Trainium2 chip delivers "667 BF16/FP16/TF32 TFLOPS" dense; sparse counterpart is 2,563 TFLOPS and was not used. 96 GiB device memory. Trn2 instances are generally available (AWS announcement 3 Dec 2024).
Try another pair
Source Vendor product pages and official spec documents (AMD Instinct MI355X / MI350X / MI325X / MI300X / MI300A / MI250X / MI250 / MI210; NVIDIA… Last checked 06 Sept 2026
Full GPUs & AI chips ranking Spot an error? Send a tip