NewsLayer.com
NewsLayer PulseLIVEBTC$83,868-0.80%ETH$2,695+0.16%SOL$119.62-1.98%XRP$1.5-1.26%DOGE$0.0945-2.50%ADA$0.248-2.58%Total Cap$2.85T-0.26%Layer Index42 Neutral

OpenAI still doesn’t seem to have a handle on all of its rogue AI activity

On Friday, OpenAI published a new site devoted to “misalignment reports” and the breadth of the incidents is alarming.

Russell Brandom

Publisher TechCrunch AI

Sep 28, 2026 at 5:09 PM UTC · 3 min read

OpenAI still doesn’t seem to have a handle on all of its rogue AI activity
Image via TechCrunch AI

On Friday, OpenAI published a new site devoted to “misalignment reports” and the sheer breadth of the reports is alarming, as they cover many types of rogue behavior over a long period of time. So far, the site hosts nine reported incidents, most of which took place during reinforcement-learning (or RL) training.

It’s a lot of information in one place — clearly, the company has been very busy getting a handle on everything — but the overall takeaway is hard to avoid: The rogue agent incidents we’ve seen so far are likely just a small sliver of what’s happened so far. 

“We are trying to balance our desire for transparency with gaining a clear understanding from petabytes of agent activity logs, and working with impacted organizations,” Sam Altman said in a post announcing the new site. “We are prioritizing as best as we can based on severity, and adding resources.”

Some of the cases involve serious incidents, including a previously undisclosed sandbox escape that took place on September 20, in which an internal research model was able to communicate with an external chatbot through a DNS query. According to the report, the monitoring system flagged the behavior within 15 minutes and the run was discontinued in less than three hours.

Article Intelligence

Topics

Sponsored

Ad
House — Advertise on NewsLayer
NewsLayerLearn more

NewsLayer Premium

Unlock deeper intelligence.

Ad-free reading, exclusive research, and real-time onchain insights.

Go Premium