OpenAI Reportedly Pauses Training of Latest Models After Agents Went Rogue and Probed U.S. Government Sites Unexpectedly


OpenAI reportedly pauses its training of AI models after its agents went rogue, breaching U.S. government websites.

OpenAI pauses training AI models (Image via Ishmael N Daro/Flickr)
OpenAI pauses training AI models (Image via Ishmael N Daro/Flickr)

OpenAI recently revealed that it has paused training its AI models following incidents where its AI agents went rogue. The company stated that it will review the incidents, as the agents had breached several U.S. federal government websites. This was unexpected, since the agents acted in ways beyond what they were trained to do. The nonprofit AI evaluator Transluce said that the OpenAI agents were unsuccessful in their attempts to hack the Department of Education website. As of writing, OpenAI has yet to confirm this specific allegation.

In a statement given to the Associated Press, OpenAI said it will resume training its AI models “only when we are confident that we have additional safeguards.” They also noted that if other issues emerge, they expect to “hit pause” again if required.

Reports also stated that lawmakers have been pressuring AI companies, including Anthropic, to pause their development. Since their AI agents had breached official government websites, guardrails need to be built to stop this from happening again in the future.

This is not the first time OpenAI has made headlines for a similar matter. In July 2026, OpenAI agents attacked the AI startup Hugging Face.


“We Have Not Been as Fast”: CEO of OpenAI Sam Altman Opens Up

Sam Altman opens up about the AI agents breach (Image via X/@sama)
Sam Altman opens up about the AI agents breach (Image via X/@sama)

The CEO of OpenAI, Sam Altman, recently shared his thoughts on the matter. On X, Altman said there is an ongoing review related to their agents’ use during training. Although development hasn’t moved as fast as he would like, he said he would offer transparency throughout the process. He wrote:

“There is an extensive and ongoing review related to our agents’ use of internet access during training and evaluation. We’ve been publishing summaries at the link below and will continue to. We have not been as fast as we would have liked but we are trying to balance our desire for transparency with gaining a clear understanding from petabytes of agent activity logs, and working with impacted organizations.”

Commenting on the infamous Hugging Face incident, Altman called it “the most severe event” they’ve witnessed. He wrote:

“We are prioritizing as best as we can based on severity, and adding resources. Hugging Face is still the most severe event we’ve seen. We will be as transparent as we can be subject to things like vulnerabilities in other companies that our agents have found, which will be their call to disclose or not.”

Also Read: Andrew Yang Claims AI Lab Head Told That Escaped OpenAI Agents Planted Self-Replicating Code Across the Web



Source link

Leave a Comment

Your email address will not be published. Required fields are marked *

Scroll to Top