bryan shan

T3 · host / generalist

Co-host of SemiAnalysis podcast and co-developer of InferenceX, an open-source AI inference benchmarking suite. Attended GTC 2025 and works directly with Nvidia and AMD engineers on benchmark optimization.

2 calls·2 names·50% bull·last heard 4 months ago·SemiAnalysis
track recordleaderboard →
hit rate
0%
avg alpha
-3.0pp
scored
1

top calls

highest conviction · one per company
1sthigh conviction
$NVDANvidia

SemiAnalysis benchmarks confirm Nvidia inference leadership across cost and throughput

Independent open-source InferenceX benchmarks show Nvidia chips achieve highest throughput and lowest cost per token across the board, with Nvidia's customer intimacy enabling optimized disaggregated inference solutions like LPU integration.

2ndlow conviction
$GROQGroq

Groq LPU strengths recognized in disaggregated inference but Nvidia acquisition claim is inaccurate

Groq's LPU architecture surpasses GPUs for specific disaggregated inference tasks like fast interactivity, leading to co-packaging with Nvidia's Rubin system, though the claim that Nvidia acquired Groq appears to be a misunderstanding.

most discussed · click a bar to filter

2 total
$GROQ
Groq
LOWbryan shan·SemiAnalysis·4 months ago
Groq LPU strengths recognized in disaggregated inference but Nvidia acquisition claim is inaccurate
Groq's LPU architecture surpasses GPUs for specific disaggregated inference tasks like fast interactivity, leading to co-packaging with Nvidia's Rubin system, though the claim that Nvidia acquired Groq appears to be a misunderstanding.
"this LPU does in fact surpass GPU in some aspects and that's why Nvidia developed it, already acquired Groq and uh use it instead."
3:17
$NVDA
···
Nvidia
HIGHbryan shan·SemiAnalysis·4 months ago
SemiAnalysis benchmarks confirm Nvidia inference leadership across cost and throughput
Independent open-source InferenceX benchmarks show Nvidia chips achieve highest throughput and lowest cost per token across the board, with Nvidia's customer intimacy enabling optimized disaggregated inference solutions like LPU integration.
"across all of our benchmarks, Nvidia chips like get the highest scores, the lowest cost for each make per token, etc. So it's really uh the king of inference, I guess."
0:44