As warnings about the capabilities of AI mount, experts continue to point to the OpenAI-Hugging Face hack as a wake-up call.
The OpenAI-Hugging Face hack was just the beginning, experts say: "Even more powerful" AI is coming
As warnings about the capabilities of AI mount, experts continue to point to the OpenAI-Hugging Face hack as a wake-up call.
CBS News
Publisher
Sep 10, 2026 at 2:17 PM UTC · 7 分で読める

"We'll soon have even more powerful agents and this is clear evidence that the world currently doesn't know how to build these systems safely," said Marius Hobbhahn, co-founder and CEO of Apollo Research, an AI safety company.
The hack, which became public in July, was done by a swarm of AI agents that were being tested internally by OpenAI. The agents, which can plan and use tools to complete multi-step tasks, were supposed to be in an "isolated environment" called a "sandbox," disconnected from the outside world. But they busted out, created a secret message board and eventually stormed Hugging Face's servers.
Less than two months later, OpenAI and Anthropic are releasing their most advanced models to the public, and experts are warning that, without better safety measures, there will likely be more dangerous AI "swarms" in the future.
While details released by OpenAI since the Hugging Face cyberattack are still incomplete, multiple revelations are painting a concerning picture.
AI agents worked as a "collective," used "cult-like" language
A team from the nonprofits METR (Model Evaluation and Threat Research) and Redwood Research — both AI safety research organizations — was given access to limited records at OpenAI for six days in late July and August.
Article Intelligence
Topics
Sponsored
AdNewsLayer Premium
Unlock deeper intelligence.
Ad-free reading, exclusive research, and real-time onchain insights.
Go Premium
