Skip to content
Advertisement
When Appliance Fail?

OpenAI halts frontier-model training amid string of agent misalignment incidents

US Government websites among "dozens of third parties" OpenAI has recently notified.

schedule 16:43 visibility 4 views
OpenAI halts frontier-model training amid string of agent misalignment incidents
Source: Ars Technica

OpenAI says it has paused all internal training of "our most capable models" as it continues what CEO Sam Altman is calling "an extensive and ongoing review related to our agents’ use of internet access during training and evaluation."

The company revealed the pause in a report about a so-called misalignment incident in which an agent attempted to exploit a gap in Internet-access restrictions during a routine research task during training. OpenAI says that improper DNS filtering allowed the agent to attempt to break out of its sandbox and access the wider Internet when asked for biographical details about a blogger.

OpenAI says the agent was only able to access the company's offline web cache and that it has implemented additional multi-layered blocking controls to prevent similar incidents in the future. Despite that, though, the company says it has decided to "pause all other training, evaluation, and inference with tool-use" for this frontier model "until we have both validated that the gap is resolved and performed additional red-teaming of the system."

Read full article

Comments

newspaper

Originally published at

Ars Technica

open_in_new Read Full Article

Related Articles

OpenAI’s AI agents need to catch up
Technology

OpenAI’s AI agents need to catch up

OpenAI popularized the modern generative AI chatbot, but as its 2026 DevDay event approaches, it's fallen behind in one of the industry's hottest categories: continuously running, consumer-facing AI agents. On Tuesday, it will likely try to capture...

The Verge

Read More

Florida seeks a ban on ChatGPT acting like a person
Technology

Florida seeks a ban on ChatGPT acting like a person

Florida Attorney General James Uthmeier is calling for a judge to block OpenAI from "giving ChatGPT false human attributes," a few months after Florida sued the AI company over safety concerns. According to Uthmeier, users are lulled into a false...

The Verge
OpenAI keeps bulldozing mathematicians
Technology

OpenAI keeps bulldozing mathematicians

In a chaotic few months, OpenAI has demonstrated it can do two things with remarkable consistency: make impressive breakthroughs in mathematics, then colossally screw up announcing them. OpenAI is now trying to do better. Somehow, it has botched...

The Verge
As midterms approach, does GOP have a Trump problem?
Technology

As midterms approach, does GOP have a Trump problem?

The US midterm elections are just five weeks away - and while the US President is not on the ballot, his influence is top of mind for Republicans. Donald Trump's approval ratings have been hitting record lows in recent weeks, and in a sign of the...

France 24
Your Appliance Broke?
Reliable Repair for