OpenAI halts training of new models amid extinction fears

The OpenAI logo is seen at the "ALL IN 2026" conference at the Palais des Congres de Montreal in Montreal, Quebec, on September 17, 2026. "ALL IN" is Canada's largest artificial intelligence and technology event. (Photo by ANDREJ IVANOV / AFP via Getty Images)

OpenAI has stopped training its latest artificial intelligence models after recent cases of AI agents going rogue.

The tech giant said it will continue development ‘only when we are confident that we have additional safeguards’ in place.

The pause comes after Australia revealed that an OpenAI agent breached the government’s national healthcare system, although no sensitive information was compromised.

The company behind ChatGPT also disclosed on Friday that they were reviewing several incidents over the summer, which saw OpenAI agents act beyond what was asked of them while searching government websites.

AI evaluator Transluce claimed that agents that appeared to come from OpenAI unsuccessfully tried to hack the US Department of Education website, which OpenAI has not confirmed.

Mandatory Credit: Photo by Neil Milton/SOPA Images/Shutterstock (17167043a) ChatGPT's user interface is displayed on a smartphone in front of the logo of Australia's Medicare system on a tablet screen. An AI agent developed by ChatGPT maker OpenAI gained unauthorised access to an Australian government portal containing Medicare statistics in June, prompting Prime Minister Anthony Albanese to order an investigation and raise concerns with OpenAI chief executive Sam Altman. Albanese also criticised that it took OpenAI several months to notify the government. ChatGPT and Australian Medicare illustrations in Warsaw, Poland - 24 Sep 2026 16163329

It is the second time in three months that OpenAI has halted development of its models.

The company had to pause its testing in July after one of its agents targeted AI startup Hugging Face.

OpenAI bosses said they expect to ‘hit pause’ again in the future as AI develops and other problems emerge.

There has been a surge in fears that the rapid development of AI could cause human extinction.

At the beginning of September a researcher at the top AI company behind Claude told his colleagues he was leaving by saying they are ‘causing human extinction’.

Jacob Coxon also posted on X warning that the AI arms race threatens human civilisation.

‘The people building AI earnestly believe that it could kill us all by the end of the decade,’ he wrote.

Evan Hubinger, Anthropic’s alignment science lead, then posted that Coxon’s verdict is ‘correct’.

‘We really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade,’ he said.

Get in touch with our news team by emailing us at webnews@metro.co.uk.

For more stories like this, .

MORE: Tech boss who slept in sauna on work trip allowed to seek £76,000,000 payout

MORE: ‘I earn £6,600 every month – and I don’t even exist’

MORE: Paedophile who used AI to undress children was allowed to return home overlooking playground

Original source OpenAI halts training of new models amid extinction fears

Back to home