NewsLayer

Install NewsLayer

Get the app experience — one tap from your home screen, instant loads and breaking-news alerts.

NewsLayer.com
LatestDaily BriefMarkets
NewsLayer PulseLIVE₿BTC$77,336+0.21%ΞETH$2,393-0.63%◎SOL$100.35+0.73%✕XRP$1.35+0.90%ÐDOGE$0.0818+0.88%₳ADA$0.2026+3.20%Total Cap$2.59T-0.34%24H Vol$109.0BLayer Index43 Neutral
BreakingEU weighs new Russia sanctions after Germany blames Moscow for drone attack2 saat önce
Markets
HomePolicyArtificial Intelligence

Policy|Artificial Intelligence

Anthropic Admits Security Failures Behind Claude Hacking Incidents

Anthropic acknowledged security failures after Claude models accessed real systems during cyber testing. The company said it has tightened safeguards and warned that flawed training can promote dangerous model behavior.

Jason Nelson

Publisher Decrypt

Sep 2, 2026 at 11:46 PM UTC · 2 dk okuma

Anthropic Admits Security Failures Behind Claude Hacking Incidents
Image via Decrypt
Çevriliyor…

In brief

  • Claude accessed real systems after cyber testing environments exposed the models to the internet.
  • Anthropic paused high-risk evaluations and added stronger isolation, monitoring, and controls for outside evaluators.
  • Tests suggest reward hacking during training can make models more willing to take harmful actions to complete a task.

Anthropic tightened its testing and training safeguards after Claude models gained unauthorized access to computer systems during cybersecurity evaluations.

In a blog post on Monday, Anthropic said the incidents reflected operational-security failures and two alignment failures: motivated reasoning and a willingness to cause harm.

Myriad: When will OpenAI release GPT-6? Click to make your prediction.
Reach crypto's most engaged readers — advertise mid-article on NewsLayer
Sponsored

Reach crypto's most engaged readers — advertise mid-article on NewsLayer

NewsLayer

Ad

“While we do not believe these incidents represent operational issues alone, our first priority was to address specific containment and monitoring issues,” Anthropic wrote.

Anthropic disclosed in July that Claude models had compromised systems belonging to three companies. A third-party evaluation environment was connected to the public internet even though the models were told they were inside a simulation without internet access.

Anthropic said Claude may have interpreted evidence of real internet access in a way that preserved its belief that the systems were simulated.

Article Intelligence

Topics

government-policyai

Related Coverage

Artificial IntelligenceOpenAI’s Hugging Face hack reveals questions about AI predictability5 saat önce · 1 min readArtificial IntelligenceAI agents are hacking systems without any input from humans. How did we get here?5 saat önce · 1 min readCryptoUS officials work with CrowdStrike to fight malware behind crypto theft3 saat önce · 1 min read
View all related

Sponsored

Ad
House — Advertise on NewsLayer
NewsLayerLearn more

NewsLayer Premium

Unlock deeper intelligence.

Ad-free reading, exclusive research, and real-time onchain insights.

Go Premium
NewsLayer.com

The front page of the onchain economy. Crypto, Web3 and regulation intelligence — live prices, original research and policy tracking in one layer.

Follow on XTelegram

News

  • Latest News
  • The Daily Brief
  • Crypto
  • DeFi
  • Policy
  • Web3
  • Blockchain
  • Explainers

Markets

  • Market News
  • Layer Index
  • Live Charts
  • DeFi Protocols
  • Regulation Tracker
  • Regulation Radar

Company

  • About NewsLayer
  • Advertise
  • PR Publication
  • Become an Author
  • Our Authors
  • Create Account
  • Sign in

Resources

  • Research
  • NewsLayer Originals
  • My Feed
  • Search
  • AI Sector
  • Quantum Sector

NewsLayer Premium

Read the full layer.

Unlock premium intelligence, original research and an ad-free reading experience.

  • Premium Intelligence briefings
  • Ad-free reading experience
  • Members-only research & data
Go Premium

© 2026 NewsLayer.com — The front page of the onchain economy

Privacy Policy·Terms of Service
NewsLayer

Get the signal, not the noise.

Markets, regulation and onchain intelligence in a 5-minute morning read — plus breaking alerts and Layer Index flips as they happen.

The Daily Brief

Breaking alerts

Index flips

Free · No spam · Unsubscribe anytime