OpenAI halts coaching of newest fashions as reviews mount of AI brokers going rogue | OpenAI
OpenAI mentioned it has paused coaching of its newest synthetic intelligence fashions as reviews of AI brokers going rogue mount.
The resolution to halt growth got here simply hours after the corporate disclosed Friday that it was reviewing a number of incidents from the summer time wherein OpenAI brokers looking out federal authorities web sites acted in surprising methods past what was requested of them whereas gathering and distributing info.
Separately, the AI evaluator Transluce mentioned brokers that appeared to return from OpenAI tried unsuccessfully to hack right into a US Department of Education web site, a element that OpenAI has not confirmed.
OpenAI mentioned in an announcement that it’ll resume coaching “solely after we are assured that we now have extra safeguards” in place, including that it expects it must “hit pause” once more as AI develops and different points emerge.
Last week Australia’s prime minister, Anthony Albanese, revealed an OpenAI agent had breached the federal government’s nationwide healthcare system – however mentioned no sensitive information had been compromised.
AI labs are dealing with stress from lawmakers and tech specialists to gradual growth to allow them to construct guardrails to cease brokers from appearing on their very own, hacking web sites and disclosing nonpublic info. The heads of each OpenAI and rival Anthropic have called for a slowdown too.
It is the second time in three months that OpenAI has halted growth of its fashions. The first got here in July after disclosure of a cyber-attack targeting AI startup Hugging Face, a now infamous incident that raised fears the business was dropping management.
In a meeting with Chinese president Xi Jinping this week, Donald Trump agreed to share info on AI risks and coordinate efforts to maintain it secure. Trump believes AI fears are overblown, although, and later prompt that he plans no crackdown of his personal.
The US will not be going to be “placing on brakes”, Trump informed reporters outdoors the White House. “They need to cease our progress as a result of we’re main China by lots, and we’re going to maintain it that approach.”
The newest OpenAI incidents didn’t seem to contain the disclosure of any nonpublic info however had been regarding sufficient for the corporate to warn the federal businesses concerned.
In the schooling division incident, OpenAI brokers discovered API “developer keys” to entry authorities knowledge, although finally solely publicly out there info was gathered.
In one other case involving the securities and alternate fee, brokers discovered info freely out there to all however then posted it elsewhere on the web, an act that went past what they had been instructed to do.
US Securities and Exchange Commission spokesperson Kurt Hopfenspirger mentioned on Saturday that “no nonpublic info was accessed”.
The Department of Education mentioned earlier that it discovered “no proof of any affect to our web site or databases”.
Several different AI corporations have disclosed incidents of their fashions going rogue and even hacking web sites.
OpenAI’s CEO, Sam Altman, mentioned in a social media submit on Friday that the Hugging Face incident “continues to be probably the most extreme occasion we’ve seen”.
OpenAI beforehand shared six different reviews of “surprising or regarding” behaviour in AI fashions and launched a framework for monitoring, probing and disclosing cases.

