Leads AI strategy and model evaluation at Clay, a company building AI-powered go-to-market engineering tools. Oversees internal evals, model selection, and migration to open-weight models for production workloads.
no scored calls yet — needs a stated position or a categorical verdict, with a matured window vs SPY
Jeff Barr's internal evaluations at Clay show Opus 5 replicates Anthropic's public benchmarks, delivering near-Fable (GPT-5.6) performance at roughly 33% of the cost, and is replacing Sonnet 4.6 workloads as a spiritual successor.
GLM 5.2 sometimes ranks at the top of Clay's internal evaluations while costing a true fraction of even cheaper OpenAI models, making it a serious contender for frontier workloads and driving rapid migration to open-weight models.
Fable (GPT-5.6) maintains best-in-class writing quality and 'immaculate vibes' that users prefer for daily work, but Anthropic's Opus 5 and open-weight models are undercutting on price for comparable tier-two workloads.