OpenAI halts frontier-model training amid string of agent misalignment incidents



News of the training pause comes just weeks after OpenAI joined other major model makers in expressing a desire to slow down model training and development over fears of potentially “catastrophic” misalignment risks. It also comes amid new reports of models improperly probing government websites during searches for high-quality data.

In a Friday blog post, OpenAI said it had notified “dozens of third parties”—including ones “operated by governments, universities, public agencies, and other institutions”—of incidents where its models either bypassed security controls or otherwise “negatively impacted” an online service in an unintended way. A New York Times report, later confirmed by OpenAI, revealed that the websites of the US Census Bureau, Securities and Exchange Commission, and Department of Education were among those affected in these newly revealed incidents. However, no private information or sensitive server infrastructure appears to have been accessed in these cases.

“The vast majority of actions we’ve reviewed were completions of mundane research tasks, such as accessing publicly available web content to answer questions,” OpenAI said in its recent blog post. “Our investigation focuses on instances where agents interacted with third-party websites in ways that went beyond their assigned tasks or intended methods… Given the scale of the review required, and the need to verify each case, this work will take months to complete.”

OpenAI’s training pause may reflect worries about corporate liability if an overzealous agent does unintentionally cause significant harm to a third-party system. Last Thursday, Australian Prime Minister Anthony Albanese promised “legal consequences” after an incident in which an OpenAI agent accessed “non-public files” from the country’s Medicare statistics portal.

While a pause in training could hurt OpenAI’s position in the highly competitive race among frontier model makers, it could also help the company’s bottom line, at least temporarily. Leaked financial documents revealed earlier this year show OpenAI’s 2024 and 2025 revenues were dwarfed by ballooning R&D expenses associated with model training.



Source link

  • Related Posts

    With Dazzle, Marissa Mayer bets your camera roll has more info on your life than your inbox

    When former Yahoo CEO Marissa Mayer earlier this month told me she was finally ready to unveil Dazzle, the personal AI assistant that raised an $8 million seed round last…

    Continue reading
    How Much VRAM Should You Look For In A Graphics Card?

    It’s probably worth buying a more expensive card now, rather than waiting. Octavian Lazar/Shutterstock Surging AI infrastructure demand has spent the year draining the world’s memory supply, leaving…

    Continue reading

    Leave a Reply

    Your email address will not be published. Required fields are marked *

    You Missed

    Greta Lee, Amanda Seyfried Fête Jean Schlumberger Conservation Project in Paris

    Greta Lee, Amanda Seyfried Fête Jean Schlumberger Conservation Project in Paris

    With Dazzle, Marissa Mayer bets your camera roll has more info on your life than your inbox

    With Dazzle, Marissa Mayer bets your camera roll has more info on your life than your inbox

    Trump’s import bans on Canadian liquor, whey, motorcycles take effect

    Trump’s import bans on Canadian liquor, whey, motorcycles take effect

    Crate & Barrel Holdings partners with Affirm to offer flexible ways to pay

    Why didn’t George Kittle take a knee? Kyle Shanahan didn’t tell him to

    Why didn’t George Kittle take a knee? Kyle Shanahan didn’t tell him to

    GTA 5 For Switch Mod Project Shuts Down, Dev Thanks Take-Two For Its “Mercy”

    GTA 5 For Switch Mod Project Shuts Down, Dev Thanks Take-Two For Its “Mercy”