newsroom
$AMD · AMD MI355X HBM advantage may close gap with prefix caching but trails Nvidia today
now playing · $AMD
$NVDA···bullish· highbryan shan
SemiAnalysis benchmarks confirm Nvidia inference leadership across cost and throughput
Independent open-source InferenceX benchmarks show Nvidia chips achieve highest throughput and lowest cost per token across the board, with Nvidia's customer intimacy enabling opt…
$GROQneutral· lowbryan shan
Groq LPU strengths recognized in disaggregated inference but Nvidia acquisition claim is inaccurate
Groq's LPU architecture surpasses GPUs for specific disaggregated inference tasks like fast interactivity, leading to co-packaging with Nvidia's Rubin system, though the claim tha…
$AMD···neutral· mediumcameron quilici
AMD MI355X HBM advantage may close gap with prefix caching but trails Nvidia today
AMD's MI355X offers 1.5x HBM capacity versus Nvidia's B200, which could enable better prefix caching performance in multi-turn workloads, but current InferenceX benchmarks show Nv…