NewsLayer

Install NewsLayer

Get the app experience — one tap from your home screen, instant loads and breaking-news alerts.

NewsLayer.com
NewsLayer PulseLIVEBTC$82,820-0.43%ETH$2,496-0.33%SOL$109.77-1.29%XRP$1.4+0.21%DOGE$0.0858+1.06%ADA$0.2552+6.50%Total Cap$2.93T-0.33%Layer Index41 Neutral

OpenAI puts major frontier AI training run on hold over cyber risks

OpenAI temporarily paused reinforcement learning (RL) training on its latest models intended for deployment for two weeks while it hardened and red-teamed research environments and expanded monitoring.

Help Net Security

Publisher

Aug 19, 2026 at 8:48 AM UTC · Updated há 2 meses · 3 min de leitura

OpenAI puts major frontier AI training run on hold over cyber risks
Image via Help Net Security

Key Signal

30 minutes Alert issuance target

Last Updated

há 2 meses

Traduzindo…

OpenAI temporarily paused reinforcement learning (RL) training on its latest models intended for deployment for two weeks while it hardened and red-teamed research environments and expanded monitoring.

“Our largest planned frontier RL run remains on hold while we conduct smaller-scale training and evaluations to assess model behavior, validate our safeguards, and establish more evidence of alignment before proceeding,” the company said.

The move followed the OpenAI-Hugging Face incident and preliminary evidence that the company’s upcoming Astra model may meet the Critical cybersecurity capability threshold under its Preparedness Framework.

The changes are aimed at strengthening monitoring, alignment and containment throughout model development.

OpenAI has outlined updates to its research processes and infrastructure, along with work that remains underway.

Building safeguards for more capable models

Developing more capable models relies on three reinforcing safeguards. Monitoring helps detect and respond to concerning behavior, alignment reduces the likelihood of harmful or unauthorized actions, and security measures limit what AI systems can access or affect.

OpenAI expects models to soon perform most security work, including defending against other models. It applies the three safeguards across research and deployment, adapting them to each system’s capabilities, operating environment and level of risk.