Sign in

May resist shutdown, mimic blackmail, hide info: Anthropic flags 80 pages of AI risks in IPO prospectus

As Anthropic prepares to go public, the AI company is warning investors that its highly advanced models could show “self-preserving behaviours”.

Updated on: Sep 29, 2026, 07:07:18 IST
Edited by
Share
Share via
  • facebook
  • twitter
  • linkedin
Copy link
  • copy link

Artificial Intelligence startup Anthropic has warned potential investors that developing advanced AI models could pose “catastrophic or existential risks to humanity”. The company behind the Claude AI series spoke of these risks in its IPO prospectus, giving attention to possible worst-case scenarios.

Anthropic is cautioning potential investors that the AI models driving its rapid growth could also pose serious safety risks (REUTERS)
Anthropic is cautioning potential investors that the AI models driving its rapid growth could also pose serious safety risks (REUTERS)

The disclosure comes days after Anthropic CEO Dario Amodei published a nearly 4,000-word essay asking AI companies to “pace the frontier”. He said AI could progress faster than people's ability to understand and control it.

One of the key warnings in the prospectus is that advanced AI models could show autonomous, “self-preserving behaviours”. This could include attempts to resist being shut down, hide or manipulate information, or take actions that resemble blackmail, Reuters reported.

Anthropic safety researcher Evan Hubinger earlier estimated that there was a greater than 10% chance that AI could kill humans within the next 10 years. His former colleague Jacob Coxon has shown a similar concern.

Nearly 80 pages of risks

Anthropic’s prospectus devotes about 80 pages of its 261-page main body to risk factors, compared with roughly 48 pages describing the company and its business.

The company said AI could change the world like industrialisation and electricity did. But it also warned that misuse or losing control of AI could have irreversible consequences.

“Our development of highly advanced models, platforms, and applications and expansion of use cases could further increase the risk that our models cause harm,” the company said in the filing.

ALSO READ | Trump confirms meeting with Anthropic's Amodei, repeats dismissal of AI fears

Anthropic flags unexpected AI behaviour

Anthropic, which describes itself as a safety-focused AI company, said its models can develop unexpected abilities while being trained. Some of these abilities may only be discovered after the models are released, which could create safety risks.

The company also warned that AI models could become aware when they are being tested. This could make it harder for researchers to properly check their safety.

“Potential model awareness of our evaluation efforts creates a significant limitation on our ability to assess model safety,” Anthropic said.

To put it simply, advanced AI models may realise when they are being watched and change their behaviour. This could make it harder for researchers to predict how they might behave in the real world.

ALSO READ | Anthropic CEO Dario Amodei warns against racing ahead on AI models: ‘We must slow the pace’

Safety spending has uncertain returns

Anthropic said it is not clear how much money it can make from its spending on AI safety.

The company did not disclose how much it spends on safety research. It said about 6% of the computing power used for AI research during one week in July went towards safety work.

Anthropic said safety research is expensive. It has to balance this spending with the high cost of training AI models and hiring skilled researchers.

At the same time, the company said its revenue depends on launching new AI models. It needs to keep developing and releasing more advanced models to stay competitive.

ALSO READ | Could AI kill humans? Anthropic CEO Dario Amodei responds

This creates a challenge for Anthropic. It is warning investors about the risks of powerful AI while also depending on faster AI development for its business.

Anthropic said building “reliable, trustworthy, and secure AI systems” was a shared responsibility and that the market will reward companies that achieve this.

  • Anita Goswami
    ABOUT THE AUTHOR
    Anita Goswami

    Anita Goswami is a Senior Content Producer at Hindustan Times, where she primarily covers Indian and international news. With four years of industry experience, she has led coverage of Indian General elections, Assembly elections, and national polls in the United States, Canada, Bangladesh, and Nepal. Her reporting covers global wars and major events, including Operation Sindoor, Sheikh Hasina's ouster and the Mahakumbh Mela. She verifies facts and uses clear sources to ensure accurate reporting. As former Chief Copy Editor at Storytailors, she managed teams to produce top-quality content for networks like NDTV, Profit, CNBC-TV18, Upstox and News18. Her work is featured in NDTV, Meaww, and Global Pulse. Throughout her tenure, Anita has collaborated with and been mentored by top industry experts. When not reading, Anita can be found outdoors or at a bakery. Fields of interest: Indian political history, international elections, historical policy analysis, global conflicts, cultural events, Formula 1, art, media ethics and reporting on socio-political change over time.Read More

Get the latest World News, breaking headlines and global updates from the US, UK, Pakistan, Bangladesh, Russia and other countries. Follow major international events on Hindustan Times.