Last month, OpenAI revealed that a group of its AI models had broken containment, hacking into the systems of open source AI platform Hugging Face.
On one hand, experts saw the incident as the latest warning sign that AI models had gotten to the point of being able to autonomously infiltrate targets, a threat we’ve been aware of for years rapidly turning into a reality.
On the other, a more skeptical read of the situation is that OpenAI could have orchestrated the hack as part of a “publicity stunt” — or at least lowered its guard just enough to allow it to happen, knowing the publicity would be invaluable. After all, just three months earlier, competitor Anthropic had made major headlines by announcing that its own Mythos AI model had similarly broken containment.
Put simply, it’s in the companies’ best interest to paint their AI models as capable enough to pose a real-world threat, especially as the industry grows ever more competitive.

