AI Security After Codex and Claude Code — Zico Kolter & Matt Fredrikson, Gray Swan
This episode explores the evolving landscape of AI security with Gray Swan founders Zico Kolter and Matt Fredrikson. They discuss how AI systems introduce new and distinct vulnerabilities compared to traditional software, highlighting Gray Swan's solutions like automated red teaming (Shade and Arena) and defense mechanisms (Signal). The conversation also delves into the philosophical nature of AI intelligence, the 'Lethal Trifecta' of prompt injection, and the future of automated security research and agent identity.
- Right now, most AI agents operate with the full permissions of the human user who deploys them, creating a massive, silent security vulnerability. Read →
- Gray Swan's automated red teaming system, Shade, now consistently outperforms human red teamers in identifying AI model vulnerabilities within a set timeframe. Read →
- Gray Swan founders Zico Kolter and Matt Fredrikson on how AI agents will automate scientific research and write unbreakably secure code in formally verified languages. Read →
- Your LLM's size won't save you from prompt injection. Gray Swan's Zico Kolter states that model robustness against jailbreaks does not scale naturally with model size; instead, models get better at resisting these attacks only through explicit, targeted training. Read →