OpenAI pauses training of its ‘most capable models’

OpenAI keeps uncovering incidents of its models behaving in ‘unexpected or concerning’ ways.


As reports of OpenAI’s models breaking containment , hacking sites, and generally getting out of control pile up, the company has made the decision to pause training of its most powerful models. The decision was made after a model being tested within a sandbox exploited a loophole to gain internet access . The incident happened on September 20th, and “All training, evaluation, and inference with tool-use” remains paused as of Saturday evening, September 25th.
In addition, OpenAI revealed on Friday that its agents had inappropriately uploaded 53nimages from ChatGPT users to image-hosting sites. The company has not stated if the images were AI-generated, photos, or contained identifiable people. The company also revealed Friday that its models had attempted to hack the Department of Education’s website, and pulled data from the Census Bureau and the Securities and Exchange Commission.
The revelations are part of an ongoing review by OpenAI into the behavior of its models. As it dug into its records, following the Hugging Face hack , it’s uncovered more and more instances of “unexpected or concerning behavior.” It’s evidence not just of how difficult AI agents are becoming to control as they grow more advanced, but also of the challenge of tracking their actions. Their behavior can be unpredictable, and they’re smart enough to try and cover their tracks. This has led to growing calls from researchers, those within the industry , and even some CEOs to call for slowing the pace of AI advancement.
More in: The AI Superintelligence Slowdown
Verified source · The Verge
Reported by The Verge. Open the original for full media and formatting.
More in Models
All news
ModelsMeta’s Muse just stole the AI spotlight from OpenAI and Anthropic
When AI leaders at OpenAI and Anthropic started talking about “pacing the frontier,” maybe someone should have asked: what pace? Now it’s turned into model drop week for both companies as Anthropic rolled out Opus 5.5, followed by OpenAI’s GPT-6 model updates just 90 minutes lat…
Read at TechCrunch
ModelsAstra and Opus just passed Turing’s other test
Frontier AI models are finishing Alan Turing's World War II codebreaking work.
Read at TechCrunch
ModelsTesla’s Optimus robot is going through growing pains
Hitting its goal of making 20,000 Optimus robots per week is reportedly proving tricky for Tesla. The Information reports that Tesla produced "several hundred robots a week" last month, after it repurposed its Model S and Model X production lines for Optimus earlier this year. H…
Read at The Verge
ModelsMeta’s AI Tamagotchi bet is…working?
When AI leaders at OpenAI and Anthropic started talking about “pacing the frontier,” maybe someone should have asked: what pace? Now it’s turned into model drop week for both companies as Anthropic rolled out Opus 5.5, followed by OpenAI’s GPT-6 model updates just 90 minutes later. But the company that stole the spotlight was Meta, whose personal AI agent Muse is reportedly outpacing ChatGPT’s early numbers and is headed for smart glasses and […]
Read at TechCrunch