TickerTain
TickerTain
NewsroomShortsPortfolioConvergence
NewsroomShortsPortfolioConvergence
←
▶ 35:36 · $CHARACTER-AI · Character AI adopts cutting-edge VLM features at hundreds of GPUs scale
episode briefing
a16z

Inferact: Building the Infrastructure That Runs Modern AI

2026-01-22 · 4 company · 4 thematic
sentiment
4 bull0 bear0 neu
speakers
wuk kuan
simon mo
quote
“Character AI saying like oh actually we already wrote it out to hundreds of GPUs at scale given just your first iteration of this feature”
—simon mo
now playing · $CHARACTER-AI
$INFERACTbullish· high· posunknown
a16z backs Inferact: VLM creators build universal inference layer
a16z invested in Inferact after granting the VLM open source project; the founders aim to make VLM the standard inference engine and build a universal runtime abstracting models a…
$NVDA···bullish· highwuk kuan
Nvidia chip roadmap drives model-hardware co-optimization imperative
Model architectures must be specialized per Nvidia chip generation (H100 vs B200 vs GB200 MVL72), creating a structural need for an abstraction layer like VLM that can optimize ac…
$AMZN···bullish· mediumsimon mo
Amazon deploys VLM at massive scale for Rufus shopping assistant
Amazon runs VLM globally to power Rufus, their front-page AI shopping assistant, validating VLM as production-ready inference infrastructure at massive scale.
$CHARACTER-AIbullish· mediumsimon mo
Character AI adopts cutting-edge VLM features at hundreds of GPUs scale
Character AI deployed VLM's speculative decoding feature to hundreds of GPUs while it was still a single unmerged pull request, demonstrating extreme velocity of open source adopt…