Founder and chief analyst of SemiAnalysis, an independent research firm covering semiconductors, AI infrastructure and compute economics.
MI455X uses CDNA 5 architecture with 320B transistors across 12 chiplets on 2nm/3nm via 3.5D packaging, delivering 432GB HBM4 at ~20 TB/s bandwidth — a memory advantage over Nvidia's Vera Rubin until Rubin Ultra arrives.
Google's full vertical integration — custom TPUs, proprietary models, and infrastructure — delivers the lowest cost per token, positioning it to capture both consumer and enterprise AI markets as inference costs become critical.
Hundreds of thousands of Trainium 2 already deployed in AWS data centers; Trainium 3 on TSMC N3P with 125B transistors and 144GB HBM3E unifies training/inference, with Anthropic's Claude and OpenAI's 2GW commitment proving hyperscale ASIC traction.