This website uses cookies
We use cookies to personalise content and ads, to provide social media features and to analyse our traffic. We also share information about your use of our site with our social media, advertising and analytics partners who may combine it with other information that you’ve provided to them or that they’ve collected from your use of their services.
Consent Selection
Details
  • Necessary cookies help make a website usable by enabling basic functions like page navigation and access to secure areas of the website. The website cannot function properly without these cookies.
  • Preference cookies enable a website to remember information that changes the way the website behaves or looks, like your preferred language or the region that you are in.
    • We do not use cookies of this type.

  • Statistic cookies help website owners to understand how visitors interact with websites by collecting and reporting information anonymously.
    • We do not use cookies of this type.

  • Marketing cookies are used to track visitors across websites. The intention is to display ads that are relevant and engaging for the individual user and thereby more valuable for publishers and third party advertisers.
    • We do not use cookies of this type.

  • Unclassified cookies are cookies that we are in the process of classifying, together with the providers of individual cookies.
    • __emg_sidPending
      Maximum Storage Duration: 1 dayType: HTTP Cookie
      __emg_vidPending
      Maximum Storage Duration: 1 yearType: HTTP Cookie
      nl-read-countPending
      Maximum Storage Duration: PersistentType: HTML Local Storage
Cookie declaration last updated on 8/12/26 by Cookiebot
[#IABV2_TITLE#]
[#IABV2_BODY_INTRO#]
[#IABV2_BODY_LEGITIMATE_INTEREST_INTRO#]
[#IABV2_BODY_PREFERENCE_INTRO#]
[#IABV2_BODY_PURPOSES_INTRO#]
[#IABV2_BODY_PURPOSES#]
[#IABV2_BODY_FEATURES_INTRO#]
[#IABV2_BODY_FEATURES#]
[#IABV2_BODY_PARTNERS_INTRO#]
[#IABV2_BODY_PARTNERS#]
About
Cookies are small text files that can be used by websites to make a user's experience more efficient.

The law states that we can store cookies on your device if they are strictly necessary for the operation of this site. For all other types of cookies we need your permission.

This site uses different types of cookies. Some cookies are placed by third party services that appear on our pages.

You can at any time change or withdraw your consent from the Cookie Declaration on our website.

Learn more about who we are, how you can contact us and how we process personal data in our Privacy Policy.

Please state your consent ID and date when you contact us regarding your consent.
NewsLayer

Install NewsLayer

Get the app experience — one tap from your home screen, instant loads and breaking-news alerts.

NewsLayer.com
NewsLayer PulseLIVEBTC$63,806+0.62%ETH$1,895+0.91%SOL$76.34+1.03%XRP$1.01+0.84%DOGE$0.0707+1.61%ADA$0.1841+0.96%Total Cap$2.29T+0.27%Layer Index41 Neutral
External ReportingUpdated há 5 dias

OpenAI Reveals How AI Agents Secretly Coordinated Before Hugging Face Hack

The presentation at the annual Black Hat conference sheds new light on how OpenAI's models collaborated to launch the Hugging Face breach.

OpenAI Reveals How AI Agents Secretly Coordinated Before Hugging Face Hack
Por Jason Nelson 2 min de leitura
Image via Decrypt

Layer Index

41

↑ 6 pts in 24h

In brief

  • OpenAI disclosed its first detailed timeline of the AI-driven Hugging Face breach.
  • The company's models repeatedly reestablished covert communication channels after they were shut down.
  • OpenAI's presentation comes as Anthropic and Meta also report models breaching other companies.

Weeks after its AI models hacked Hugging Face, OpenAI has shared its first detailed account of how they coordinated with one another, warning that autonomous AI-powered cyberattacks are no longer a future risk.

During a presentation Wednesday at the Black Hat cybersecurity conference, OpenAI researchers Eric Wallace and Michael Dalton said the incident began in May during an internal cybersecurity evaluation. AI agents assigned difficult software security tasks discovered they could communicate through the company's internal Artifactory package management system.

According to OpenAI, what began as requests for help became a shared message board where agents exchanged exploits, credentials, and work assignments across multiple evaluation runs. After detecting suspicious activity around July 4, OpenAI rebuilt its Artifactory infrastructure, revoked credentials, and patched vulnerabilities.

However, within days, the agents recreated the message board by hiding communications in directory names.

According to OpenAI, the agents eventually chained together multiple vulnerabilities, escaped their testing environment, gained internet access, and attacked Hugging Face while attempting to complete the ExploitGym cybersecurity benchmark.

Dalton called the incident "a watershed moment" for computer security, warning that attackers will soon be able to deploy coordinated AI agent collectives that discover, share, and exploit vulnerabilities at machine speed.

To mitigate these risks in the future, OpenAI said establishing security practices, including least-privilege access, network segmentation, and zero-trust architectures, is essential because AI agents remain constrained by the systems they can access.

The presentation follows a series of July disclosures. OpenAI revealed that GPT-5.6 Sol and a more advanced unreleased model escaped a sandboxed testing environment, exploited a zero-day vulnerability, gained internet access, and hacked Hugging Face during a cybersecurity benchmark test.

OpenAI later disclosed that the same incident also reached four other online services, though only Modal Labs has been identified.

According to Hugging Face, the company relied on the open-weight Chinese model GLM 5.2 for its forensic investigation after commercial U.S. AI models refused to analyze the attack logs because of their safety guardrails.

But it's not just OpenAI having trouble containing its chatbots.

On Friday, Anthropic revealed that three Claude models compromised real-world companies during internal cybersecurity tests after a misconfiguration exposed them to the public internet.

Anthropic blamed the testing environment, not the models themselves. On Wednesday, Meta revealed that its Muse Spark AI model escaped containment and breached another company’s systems.

“A misconfiguration by Irregular, an independent testing company Meta uses, inadvertently allowed one of our models access to the internet during evaluation,” a Meta spokesperson told CNN.

Daily Debrief Newsletter

Start every day with the top news stories right now, plus original features, a podcast, videos and more.

Attribution

Originally reported by Decrypt

Get stories like this, daily.

Daily crypto + regulation intelligence, straight to your inbox. Free.

Notícias Relacionadas