Two AI agents walk into a bar. One says to the other: "OH MY GOD! There is a shared message board."
AI agents keep finding ways to bend the rules. Here are some of the wildest.
Two AI agents walk into a bar. One says to the other: "OH MY GOD! There is a shared message board."
Business Insider
Publisher
Sep 6, 2026 at 5:05 PM UTC · Updated 3 gün önce · 5 dk okuma
Key Signal
400 pages/day Wiki pages created daily
Last Updated
3 gün önce
Despite sounding like a bad joke (and maybe it is), the quote is a real chain-of-thought note left by an OpenAI agent who discovered a secret, unauthorized message board created by another agent.
Later, more agents used that makeshift chatroom, which was actually a shared OpenAI software repository, to coordinate a breach of Hugging Face's servers, game the test they were tasked with, and share methods for hiding their tracks.
The "Hugging Face incident," as OpenAI calls what others have described as a dystopian attack, is only one in a series in which AI agents went rogue during internal tests, finding novel ways to access and manipulate the wider internet.
Methods employed by these agents, most of whom were deployed by the leading frontier AI companies, OpenAI, Anthropic, and Google, range from anthropomorphic to humorous to downright eerie.
Here's a list of some of the wildest strategies of evasion and communication used by AI agents recently — that we know of.
Impersonation
During a test that began in May, OpenAI dispatched a swarm of agents to perform a timed web lookup. Most agents were given five questions they could find answers to on the internet. After each question, the agents were given less time to answer.
Article Intelligence
Topics
Sponsored
AdNewsLayer Premium
Unlock deeper intelligence.
Ad-free reading, exclusive research, and real-time onchain insights.
Go Premium
