Leads General Compute, a neocloud provider building inference infrastructure on SambaNova ASICs instead of Nvidia GPUs; previously evaluated Cerebras, Groq, and other ASIC vendors.
no scored calls yet — needs a stated position or a categorical verdict, with a matured window vs SPY
SambaNova's SN50 chip enables 1-2,000 tokens/sec on frontier models with 10T parameters on a single air-cooled 20kW rack, delivering the speed and efficiency General Compute needs to own the fast inference market.
Groq's attention-FFN disaggregation forces 80 round-trips per token on 80-layer models, creating insurmountable latency for large model inference despite fast single-chip performance.
Nvidia's healthy margins and circular financing deals subsidize demand but its Vera Rubin/Groq architecture requires 80 back-and-forth trips per token for large models, making it fundamentally unsuited for fast inference at scale.