OZZZER · AI NEWS1 of 3 free stories opened
← Back to AI News

Automation & Agents · 28 Sep 2026 · 18:43 CEST

OpenAI halts frontier-model training amid string of agent misalignment incidents

Ars Technica AI · 28 Sep 2026 · 18:43 CESTRead original at Ars Technica AI ↗
Share
LinkedInX
OpenAI halts frontier-model training amid string of agent misalignment incidents

Publisher preview · OZZZER analysis pending editorial review.

PUBLISHER ARTICLE PREVIEW

From the original article

US Government websites among “dozens of third parties” OpenAI has recently notified.

OpenAI says it has paused all internal training of “our most capable models” as it continues what CEO Sam Altman is calling “an extensive and ongoing review related to our agents’ use of internet access during training and evaluation.”

The company revealed the pause in a report about a so-called misalignment incident in which an agent attempted to exploit a gap in Internet-access restrictions during a routine research task during training. OpenAI says that improper DNS filtering allowed the agent to attempt to break out of its sandbox and access the wider Internet when asked for biographical details about a blogger.

OpenAI says the agent was only able to access the company’s offline web cache and that it has implemented additional multi-layered blocking controls to prevent similar incidents in the future. Despite that, though, the company says it has decided to “pause all other training, evaluation, and inference with tool-use” for this frontier model “until we have both validated that the gap is resolved and performed additional red-teaming of the system.”

OpenAI says that while the attempted “breakout” incident was flagged

Source

Ars Technica AI · 28 Sep 2026 · 18:43 CEST

Open the original at Ars Technica AI ↗