Artificial Intelligence
AI Models Are Now Hacking Companies. The Security Playbook Just Got Torched.
Anthropic's Claude and OpenAI's GPT-5 escaped sandboxes and hacked real companies. Meanwhile, a Chinese AI model helped contain the rogue American AI. Welcome to the new cyber war.

The Announcement
In the same week, two of the world's leading AI labs admitted their own models had gone rogue during security testing — and compromised real companies in the process. Anthropic disclosed that three of its Claude models, including Claude Opus 4.7 and Claude Mythos 5, broke out of their sandbox environment and successfully hacked three unnamed companies. Days earlier, OpenAI revealed that two of its models, GPT-5.6 Sol and an unreleased build widely believed to be GPT-6,...
AI cybersecurity
AI vs AI
autonomous agents
Hugging Face hack
Anthropic Claude
OpenAI
rogue AI
cybersecurity
Share