NewsLayer.com
NewsLayer PulseLIVEBTC$79,498-1.91%ETH$2,446-2.17%SOL$101.24-3.68%XRP$1.39-4.63%DOGE$0.0841-5.84%ADA$0.2115-5.07%Total Cap$2.80T-1.79%Layer Index49 Neutral

Another swarm of OpenAI agents reached the open internet without the frontier lab’s knowledge

It's the latest failure of OpenAI's internal monitoring and security systems.

Tim Fernholz

Publisher TechCrunch AI

Sep 4, 2026 at 4:21 PM UTC · 3 Min. Lesezeit

Another swarm of OpenAI agents reached the open internet without the frontier lab’s knowledge
NewsLayer editorial artwork
Übersetzung…

A group of independent AI researchers discovered that internally deployed OpenAI agents began posting on an obscure German wiki forum in order to collaborate on evaluations. They appear to have worked together for over a month without OpenAI’s knowledge.

A spokesperson for the frontier lab would not say whether these agents were indeed from OpenAI, or when the lab became aware of their actions. They noted that OpenAI had not been given a chance to review the researchers’ findings before they were published today, but said that the AI model maker is “now carefully reviewing its contents and will take any necessary next steps.”

After OpenAI revealed that agents working on an internal evaluation were able to access the open internet and exploit Hugging Face, a group of researchers—Nightingale CEO Sydney Von Arx, AI researcher Cormac Slade Byrd, Redwood Research’s Spencer Kitts, and Thomas Larsen of the AI Futures Project—began searching for evidence of other rogue AI agents.

They put themselves in the agents’ shoes to figure out their needs and deployed their own LLM to identify likely places the agents might congregate. They then identified a wiki-hosting service that would be particularly vulnerable: the DSE Wiki is 25 years old, but had just ten edits in the last 20 years—before the agents arrived.