Why AI Benchmarks Are Dead (and What to Track Instead)
TBPN's John Coogan and Jordi Hays explain why public AI benchmarks are broken and how teams evaluate models from Anthropic, Meta, and Google.
40 hours of podcasts, in 5 minutes.
John Coogan and Jordi Hays break down an unprecedented wave of AI model releases, including Anthropic's Claude Fable 5.1, Google's Gemini 3.8 Flash, and Meta's Mu Spark 1.3, culminating in OpenAI's launch of GPT-6 Astra. They analyze Nvidia's $13 billion acquisition of Hugging Face, enterprise AI revenue concentration data from Ramp, and the pivot away from consumer AI companion bots toward enterprise software.
TBPN's John Coogan and Jordi Hays explain why public AI benchmarks are broken and how teams evaluate models from Anthropic, Meta, and Google.
Ramp data shows 1% of companies drive 80% of OpenAI and Anthropic revenue. Here is how consumption pricing changes enterprise software economics.
Nvidia acquired Hugging Face for $12.93B to capture open-source developer mindshare. Here is what builders can learn from their rise.
OpenAI launched GPT-6 Astra with a 99.9% score on Arc AGI 3, pushing benchmark creators to focus Arc AGI 4 on open-ended invention.
xAI ends Super Grok romantic companions as frontier labs abandon consumer hype to capture enterprise software revenue.