newsroom
Semiconductors · Huawei Ascend 910 and Alibaba chips become inference targets for Chinese models
now playing · Semiconductors
Quantization breakthroughs enable 100-10,000x compute efficiency gains within 3 years
Ternary/binary quantization (1.58 bits) with <5% accuracy loss enables frontier models on smartphones. Sub-1-bit quantization via sparsity and low-rank factorization going mainstream in 12…
Kimi K3 and Inkling trigger open-source frontier commoditization
Chinese lab Moonshot's 2.8T parameter Kimi K3 matches GPT-4.5-class performance on the cost-performance Pareto frontier and will release weights open-source. Combined with Murati's Inkling,…
Chinese open-weight models commoditize frontier intelligence, destroying closed-lab moats
Kimi K3 proves open models can match frontier performance at 1% of training cost. Frontier intelligence is now a perishable asset with weeks-long shelf life. Value shifts from model weights…
AI Infrastructuretailwindscore 9/10imad
Quantization breakthroughs (ternary, binary, sub-1-bit) to 100x inference efficiency
Prism ML's Bonsai 27B runs at ternary (1.58-bit) with 5% accuracy drop; Tencent's binary compression and Samsung's sub-1-bit 'nano quant' prove the trajectory. Imad predicts sub-1-bit mains…
Jevons paradox drives chip demand higher as software efficiency lowers compute cost
100x software efficiency gains don't reduce silicon demand—they expand total compute consumption. Nvidia Blackwell/Rubin sold out for years. Inference providers (Fireworks, Modal, Base10) r…
Every corporation must evaluate self-hosted open models now or lose competitive advantage permanently
Kimi K3 and Inkling enable internal fine-tuning on proprietary data without vendor lock-in. Companies need crash programs with best advisors to decide build vs buy. Inference providers (Fir…
Model-swapping architecture becomes the new enterprise moat
Frontier intelligence is now a perishable asset (shelf life: weeks). Value shifts to 'interfaces' — orchestration layers that hot-swap models (Kimi, Inkling, Fable, proprietary) per task. E…
Frontier lab valuations face 75% compression; value shifts to reliability and deployment layer
Government review delays cut value 50%, open-weight parity cuts another 50%. Frontier labs become one ingredient; competitive moat moves to integrated tools, security, ease of deployment. J…
Cybersecuritytailwindscore 8/10immad mostaque
AI cyber offense proliferation makes defense the biggest VC category; 1-2 quarter lag for open models
Frontier models (Fable, Opus) trained on CVE data enable cyber weapons. Chinese models lack this data but can be fine-tuned. Open-source cyber-capable models emerging in 1-2 quarters. Every…
Gavin Baker thesis: all non-frontier-lab stocks benefit from model commoditization
As foundation model layer commoditizes (open weights, 100x cheaper inference), every application-layer company, enterprise, and inference provider gains. Only pure-play frontier labs (OpenA…
US chip embargo backfired, accelerating Chinese algorithmic efficiency
Export controls on H800s/Blackwell forced Chinese labs (Moonshot, DeepSeek) to innovate on data curation, Muon optimizer, linear attention, and quantization — achieving frontier performance…
US chip embargo backfired, accelerating Chinese efficiency innovation without stopping progress
H800 embargo was 'worst case scenario'—enough to irritate but not stop China. Forced quantization, mixture-of-experts, and data curation innovations that are now permanent. China now full-s…
US immigration failure cedes AI talent to China; 70% of elite researchers non-US citizens
70% of elite AI researchers are Chinese, Indian, Taiwanese, UK. Stapling green cards to PhDs is zero-friction win. China's ecosystem now thriving—80% of Chinese PhDs return vs Indians stayi…
China scaling humanoid robots to millions; West needs economic competition not combat demos
150 Chinese humanoid companies, Unitree at 11k units heading to 11M/year. MMA combat demos stress-test engineering but set dangerous inductive priors. West (ProRL) needs economically produc…
Orbital data centers inevitable by early 2030s; SpaceX Starship reuse economics cross over terrestrial
Starship rapid reuse demonstrated. SpaceX AI (xAI) partnered with Anthropic for Colossus orbital compute. Unit economics cross over by early 2030s. OpenAI dismissing space now but will pivo…
FINRA-style SEC regulator for frontier AI could block Chinese open weights in US
The administration is exploring a FINRA-like self-regulatory organization under the SEC to govern frontier model deployment. This could make it economically infeasible for US public corpora…
Cybersecuritytailwindscore 7/10imad
Cyber defense becomes largest VC category as open models proliferate offensive capabilities
Open-weight models (Kimi K3, future Chinese releases) will gain cyberattack capability within 1-2 quarters via fine-tuning on CVE data. Every company needs AI-native cyber defense. Imad: 'm…
US immigration failure cedes 70% of elite AI researchers to China/India/UK
70% of top AI researchers are non-US citizens (Chinese, Indian, Taiwanese, UK). Stapling green cards to PhDs is the lowest-friction fix. Moonshot founder Yang Zhilin started his first AI st…
AI superforecasters (Cassie) to plug into capital markets, validating EMH
AI models now match human superforecasters on ForecastBench. When hyper-forecasting connects to algo trading (already dominant by volume), markets will pre-price human collective actions fa…
China's humanoid robot push (Unitree, Engine AI) creates military and commercial urgency for West
150+ Chinese humanoid companies, state-backed, staging MMA combat demos (Engine AI T800s punch 4x Tyson). Alex warns PLA infantry integration is likely. West needs competitive humanoid prog…
Orbital data centers cross economic viability by early 2030s per SpaceX-Anthropic deal
Starship launch cadence and Colossus 2-scale compute make space-based data centers inevitable. Anthropic's partnership with SpaceX AI for orbital compute validates the thesis. Sam Altman's…
Semiconductorstailwindscore 6/10imad
Huawei Ascend 910 and Alibaba chips become inference targets for Chinese models
Kimi K3 optimized for static shapes on Huawei/Alibaba silicon (64-node clusters). While US inference providers (Modal, Fireworks) will serve same models 10-100x cheaper on Blackwell/Rubin,…