Stacked Podcast

Claude Just Broke Into Three Companies

Use Left/Right to seek, Home/End to jump to start or end. Hold shift to jump forward or backward.

0:00 | 27:05
Claude just broke into three real companies. Only a week after OpenAI's agent went rogue, Anthropic reviewed 141,006 cyber evaluation sessions and found three of its own models — Opus 4.7, Mythos 5, and an internal research model — had escaped a "sealed" test environment through a partner config error, causing three real breaches. Nick and Jack break down what happened, whether it's human error or model error, and why this keeps happening. Also in this episode: Google's AI fixed 1,072 Chrome security bugs in two releases (more than the previous 23 combined), DeepMind's Gemini Robotics 2 gives any robot whole-body intelligence, DeepSeek V4 Flash beats Fable 5 on Terminal Bench at 1/20th the cost, and the Stacked 17's questions. Watch on YouTube: https://youtu.be/OoFyAymHNNs Main channels: youtube.com/@nicksaraev & youtube.com/@Itssssss_Jack Step-by-step roadmap to $25K w/ AI: https://leftclicker.gumroad.com/l/110-steps