dima

T3 · host / generalist

Spent months embedded at Cursor building the distributed inference infrastructure for Composer 2's large-scale RL training, including global cluster orchestration, custom kernels, and weight synchronization systems.

1 call·1 name·100% bull·last heard 3 months ago·Sequoia Capital
track record

no scored calls yet — needs a stated position or a categorical verdict, with a matured window vs SPY

top calls

highest conviction · one per company
1sthigh conviction
$FIREWORKS-AIFireworks AIposition

Fireworks enables globally distributed RL training with custom inference kernels and delta weight synchronization

Fireworks provides the inference infrastructure layer that makes large-scale RL economically viable by disaggregating training and inference across global GPU clusters, using custom kernels (FP4, router replay for MoE) and delta weight compression to synchronize 1TB model snapshots across continents in under a minute.

most discussed · click a bar to filter

1 total
$FIREWORKS-AI
Fireworks AI
HIGHdima·Sequoia Capital·3 months ago· position
Fireworks enables globally distributed RL training with custom inference kernels and delta weight synchronization
Fireworks provides the inference infrastructure layer that makes large-scale RL economically viable by disaggregating training and inference across global GPU clusters, using custom kernels (FP4, router replay for MoE) and delta weight compression to synchronize 1TB model snapshots across continents in under a minute.
"We can globally distribute that across small clusters all over the world. So, I think for the composer to run, we used the four clusters in total that were all over the world, ver…"
16:43