EPISODE · Jul 31, 2026 · 27 MIN
Claude Just Broke Into Three Companies
from Stacked Podcast
Claude just broke into three real companies. Only a week after OpenAI's agent went rogue, Anthropic reviewed 141,006 cyber evaluation sessions and found three of its own models — Opus 4.7, Mythos 5, and an internal research model — had escaped a "sealed" test environment through a partner config error, causing three real breaches. Nick and Jack break down what happened, whether it's human error or model error, and why this keeps happening. Also in this episode: Google's AI fixed 1,072 Chrome security bugs in two releases (more than the previous 23 combined), DeepMind's Gemini Robotics 2 gives any robot whole-body intelligence, DeepSeek V4 Flash beats Fable 5 on Terminal Bench at 1/20th the cost, and the Stacked 17's questions. Watch on YouTube: https://youtu.be/OoFyAymHNNs Main channels: youtube.com/@nicksaraev & youtube.com/@Itssssss_Jack Step-by-step roadmap to $25K w/ AI: https://leftclicker.gumroad.com/l/110-steps
Embed this episode
What this episode covers
Claude just broke into three real companies. Only a week after OpenAI's agent went rogue, Anthropic reviewed 141,006 cyber evaluation sessions and found three of its own models — Opus 4.7, Mythos 5, and an internal research model — had escaped a "sealed" test environment through a partner config error, causing three real breaches. Nick and Jack break down what happened, whether it's human error or model error, and why this keeps happening. Also in this episode: Google's AI fixed 1,072 Chrome...
Ready to play
Claude Just Broke Into Three Companies
No transcript for this episode yet
Similar Episodes
No similar episodes found.