Claude AI firm warns AI may pose ‘existential risk’ to humanity

FILE PHOTO: An Anthropic logo is displayed at Moscone Center during the Dreamforce 2026 technology summit in San Francisco, California, US, September 17, 2026. REUTERS/Carlos Barria/File Photo

The company behind Claude AI has warned investors that advanced artificial intelligence could pose a ‘catastrophic or existential risk to humanity’.

Anthropic, which is one of the most valuable AI companies in the world, has highlighted risks associated with its models in its IPO prospectus.

The company said future models could exhibit ‘self-preserving behaviours’, including attempts to ‘resist shutdown’, ‘conceal or manipulate information’ or even behave in ways resembling ‘blackmail’.

‘Our development of highly advanced models, platforms, and applications and expansion of use cases could further increase the risk that our models cause harm,’ Anthropic said in the filing.

The AI firm, which has positioned itself as a safety-first company, devoted around 80 pages of its 261-page prospectus to risk factors.

This was nearly twice the 48 pages it used to describe its business.

FILE PHOTO: OpenAI and Anthropic logos are seen in this illustration taken June 11, 2026. REUTERS/Dado Ruvic/Illustration//File Photo
SHENZHEN, CHINA - SEPTEMBER 22: In this photo illustration, the app icons for ChatGPT, Meta's Muse and Anthropic's Claude are displayed in a folder labeled "Artificial Intelligence" on a smartphone screen on September 22, 2026, in Shenzhen, Guangdong Province, China. Strong demand for Meta Platforms' (NASDAQ: META) Muse AI assistant has renewed investor enthusiasm for artificial intelligence ahead of the company's September 2324 Connect conference, where it plans to showcase advances in AI, smart glasses and virtual reality. (Photo Illustration by Cheng Xin/Getty Images)

Anthropic also said in the prospectus: ‘Potential model awareness of our evaluation efforts creates a significant limitation on our ability to assess model safety.’

It also warned that models can sometimes develop unexpected capabilities during training which may not be discovered until they have been deployed.

The company said developing safe AI was ‘resource-intensive’ and required it to balance spending on computing power, talent and safety research.

‘We believe building reliable, trustworthy, and secure AI systems is a collective responsibility and that the market will reward it,’ Anthropic added.

While many public companies have outlined product risks to investors, few have issued warnings suggesting their technology could cause potential human extinction.

Anthropic safety researcher Evan Hubinger estimated that there is a greater than 10% chance AI could kill humans within the next decade, which echoed former colleague Jacob Coxon’s warning.

Dario Amodei, the CEO of Anthropic, also recently voiced his own concerns about the rate of AI development.

In a near 4,000-word essay, he said that he is concerned AI has been ‘advancing drastically faster’ since this summer.

Metro has contacted Anthropic for comment.

MORE: Bill Gates warns AI ‘could cause a billion deaths’

MORE: OpenAI halts training of new models amid extinction fears

MORE: Tech boss who slept in sauna on work trip allowed to seek £76,000,000 payout

Original source Claude AI firm warns AI may pose ‘existential risk’ to humanity

Back to home