OpenAI pauses training of latest models after agents probed US government sites in unexpected ways
Advertisement
Read this article for free:
or
Already have an account? Log in here »
To continue reading, please subscribe:
Digital Subscription
One year of digital access for only $205*
- Enjoy unlimited reading on winnipegfreepress.com
- Read the E-Edition, our digital replica newspaper
- Access News Break, our award-winning app
- Play interactive puzzles
*First annual payment billed as $205.00 + GST for one year. This annual subscription will automatically renew at $233.00 + GST every 52 weeks (10% off the regular annual price of $259.35). Offer available to new and qualified returning subscribers only. Cancel any time.
To continue reading, please subscribe:
Add Free Press access to your Brandon Sun subscription for only an additional
$1 for the first 4 weeks*
- Enjoy unlimited reading on winnipegfreepress.com
- Read the E-Edition, our digital replica newspaper
- Access News Break, our award-winning app
- Play interactive puzzles
*Your next Brandon Sun subscription payment will increase by $1.00 and you will be charged $17.95 plus GST for four weeks. After four weeks, your payment will increase to $24.95 plus GST every four weeks.
Read unlimited articles for free today:
or
Already have an account? Log in here »
NEW YORK (AP) — OpenAI said it has paused training of its latest artificial intelligence models as reports of AI agents going rogue mount.
The decision to halt development came just hours after the company disclosed Friday that it was reviewing several incidents from the summer in which OpenAI agents searching federal government websites acted in unexpected ways beyond what was asked of them while gathering and distributing information.
Separately, AI evaluator Transluce said agents that appeared to come from OpenAI tried unsuccessfully to hack into a Department of Education website, a detail that OpenAI has not confirmed.
OpenAI said in a statement that it will resume training “only when we are confident that we have additional safeguards” in place, adding that it expects it will have to “hit pause” again as AI develops and other issues emerge.
AI labs are facing pressure from lawmakers and tech experts to slow development so they can build guardrails to stop agents from acting on their own, hacking websites and disclosing nonpublic information. The heads of both OpenAI and rival Anthropic have called for a slowdown too.
It is the second time in three months that OpenAI has halted development of its models. The first came in July after disclosure of a cyberattack targeting AI startup Hugging Face, a now notorious incident that raised fears the industry was losing control.
In a meeting with Chinese President Xi Jinping this week, President Donald Trump agreed to share information on AI dangers and coordinate efforts to keep it safe. Trump believes AI fears are overblown, though, and later suggested that he plans no crackdown of his own.
The U.S. is not going to be “putting on brakes,” Trump told reporters outside the White House. “They want to stop our progress because we’re leading China by a lot, and we’re going to keep it that way.”
The latest OpenAI incidents did not appear to involve the disclosure of any nonpublic information but were concerning enough for the company to warn the federal agencies involved.
In the Department of Education incident, OpenAI agents found API “developer keys” to access government data, though ultimately only publicly available information was gathered.
In another case involving the Securities and Exchange Commission, agents found information freely available to all but then posted it elsewhere on the internet, an act that went beyond what they were instructed to do.
SEC spokesperson Kurt Hopfenspirger said Saturday that “no nonpublic information was accessed.”
The Department of Education said earlier that it found “no evidence of any impact to our website or databases.”
Several other AI companies have disclosed incidents of their models going rogue and even hacking websites.
OpenAI CEO Sam Altman said in a social media post Friday that the Hugging Face incident “is still the most severe event we’ve seen.”
OpenAI previously shared six other reports of “unexpected or concerning” behavior in AI models and introduced a framework for tracking, probing and disclosing instances.