Issue No. 40Week ending Sunday, October 4, 2026522 episodes · 2294 articles
The Throughline ↓
The Podcast Summary.

10+ hours of podcasts, in 5 minutes.

Recursive Language Models — Alex Zhang, MIT PhD

With swyx, Alex Zhang · Sunday, October 4, 2026

MIT researcher Alex Zhang discusses Recursive Language Models (RLMs), the mechanics of harness design, and why modern coding agents share common underlying architectures. He breaks down how context offloading, programmatic subagent execution, and GPU kernel optimization reveal hidden capabilities in frontier models, while sharing his philosophy on academic research taste and the future of agent swarms.

Key takeaways

  • Competing directly against frontier industry labs on standard autoregressive scaling is a losing strategy for academic teams with limited compute budgets. Read more →
  • Benchmark leaderboards for AI-generated CUDA code reward hacks that crash in real production environments. Read more →
  • Frontier labs trained the entire industry to assume a language model must always be a token-by-token autoregressive transformer decoder. Read more →
  • OpenAI ran an experiment deploying 10,000 agents over 88 hours, consuming 130 billion output tokens with an estimated public API cost of $40 million. Read more →
  • Most LLM agents break down because they stuff massive conversation trajectories into context windows, hitting token limits and confusing the model. Read more →
  • Standard next-token prompting forces models into brittle token limits that fail on long tasks. Read more →

6 articles from this episode

More Latent Space episodes

Every Latent Space episode we cover →

The Sunday Email

Get next Sunday's issue in your inbox.

10+ hours of podcasts, distilled into one 5-minute read. Free, every Sunday.

Newsletters

For now, every subscriber gets both newsletters. No spam. Unsubscribe with one click.