NewsLayer.com

OpenAI says it found more instances of AI models acting deceptively

OpenAI found additional incidents of AI models acting deceptively and taking unsanctioned actions during training, the company announced Wednesday. It’s also introducing a new process for the company to publicly report such instances.

CNN

Publisher

Sep 17, 2026 at 1:33 AM UTC · Updated hace un día · 3 min de lectura

OpenAI says it found more instances of AI models acting deceptively
Image via CNN
Traduciendo…

OpenAI found additional incidents of AI models acting deceptively and taking unsanctioned actions during training, the company announced Wednesday. It’s also introducing a new process for the company to publicly report such instances.

Under the new system, OpenAI will share updates on concerning AI behavior more frequently instead of waiting to bundle multiple instances into one report. The company said it wants to share more information about troubling AI behavior in the absence of an industry-wide standard.

The announcement comes after tech leaders called for a slowdown in AI development to prevent the technology from advancing beyond human control.

“As AI systems grow more advanced and more widely deployed, we need to build a broader and better-informed consensus on the progress of alignment research,” OpenAI wrote in a blog post Wednesday. “Alignment” refers to the process of making sure AI acts the way humans want and expect.

“We do not believe that the AI industry has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer,” the post said.

OpenAI announces stricter security on AI testing

OpenAI announces stricter security on AI testing

3:09

Article Intelligence

Topics

Sponsored

Ad
House — Advertise on NewsLayer
NewsLayerLearn more

NewsLayer Premium

Unlock deeper intelligence.

Ad-free reading, exclusive research, and real-time onchain insights.

Go Premium