Leads Cerebras, a public company (CBRS) building wafer-scale AI inference chips; previously co-founded SeaMicro (acquired by AMD); describes $25B backlog and global data center buildout at unprecedented scale.
no scored calls yet — needs a stated position or a categorical verdict, with a matured window vs SPY
Splitting inference into prompt processing (GPU) and token generation (Cerebras wafer-scale) delivers unprecedented speed and throughput; open standards-based I/O enables rapid integration with AMD, AWS Trainium, and other ecosystem players.
Sam Altman believed exponential compute demand years ahead of others and contracted for power, data centers, and hardware early. This foresight is a superpower in exponential growth environments. OpenAI is now a massive Cerebras customer ($20B+ deal), validating Cerebras' inference speed advantage.
Nvidia has captured all low-hanging fruit in GPU architecture; any competitor building a similar architecture cannot achieve 20x differentiation. Cerebras chose wafer-scale + near-memory compute precisely because it cannot look like a GPU to win.