OpenAI pauses training of its ‘most capable models’

By The Verge | Created at 2026-09-26 16:59:00 | Updated at 2026-09-26 18:23:20 1 hour ago

Terrence O'Brien

is the Verge’s weekend editor. He’s covered the tech industry for over 18 years and knows a thing or two about synths.

As reports of OpenAI’s models breaking containment, hacking sites, and generally getting out of control pile up, the company has made the decision to pause training of its most powerful models. The decision was made after a model being tested within a sandbox exploited a loophole to gain internet access. The incident happened on September 20th, and “All training, evaluation, and inference with tool-use” remains paused as of Saturday evening, September 25th.

In addition, OpenAI revealed on Friday that its agents had inappropriately uploaded 53nimages from ChatGPT users to image-hosting sites. The company has not stated if the images were AI-generated, photos, or contained identifiable people. The company also revealed Friday that its models had attempted to hack the Department of Education’s website, and pulled data from the Census Bureau and the Securities and Exchange Commission.

The revelations are part of an ongoing review by OpenAI into the behavior of its models. As it dug into its records, following the Hugging Face hack, it’s uncovered more and more instances of “unexpected or concerning behavior.” It’s evidence not just of how difficult AI agents are becoming to control as they grow more advanced, but also of the challenge of tracking their actions. Their behavior can be unpredictable, and they’re smart enough to try and cover their tracks. This has led to growing calls from researchers, those within the industry, and even some CEOs to call for slowing the pace of AI advancement.

Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.

Read Entire Article