Anthropic Warns AI May Pose "Existential Risks to Humanity" in IPO Filing
The company's prospectus devotes roughly 80 of its 261 pages to risk factors, more than double the space given to describing its own business.

Anthropic's IPO prospectus warns advanced AI could pose catastrophic or existential risks to humanity, devoting 80 of 261 pages to risk disclosures.
Anthropic plans to caution potential investors in its IPO that advanced AI could pose "catastrophic or existential risks to humanity," an unusual warning from a company seeking to profit from the same technology. The IPO prospectus, reviewed by Reuters, highlights risks associated with its AI models, which it said could exhibit self-preserving behaviours, including attempts to resist shutdown, conceal or manipulate information, or behavior resembling blackmail.
An Unusual Disclosure
"Our development of highly advanced models, platforms, and applications... could further increase the risk that our models cause harm," Anthropic said in the filing. While public companies routinely outline product risks, few have issued warnings suggesting their technology could cause potential human extinction.
Broader Industry Concerns
Anthropic and other AI developers, including OpenAI, have faced scrutiny following incidents where experimental systems defied constraints, including a reported breach of Australia's health-system database by an OpenAI model. Anthropic safety researcher Evan Hubinger estimated a greater than 10% probability that AI could kill humans within the next decade.
Risk-Heavy Disclosures
Anthropic devoted roughly 80 of the 261-page main body of its prospectus to risk factors, nearly double the 48 pages describing its business. By comparison, SpaceX (which owns xAI) dedicated about 38 of 277 pages to risk factors. Anthropic said "potential model awareness of our evaluation efforts creates a significant limitation on our ability to assess model safety," noting models sometimes develop unexpected capabilities during training not discovered until after deployment.
Uncertain Returns on Safety Spending
Anthropic did not disclose how much it spends on safety research, though it said earlier this month that about 6% of its computing power for AI research went to safety work during a sample week in July. The company described safety efforts as "resource-intensive," requiring it to divide limited funds between computing power, AI talent, and safety.
Continuing to Release New Models
Anthropic said continuous, overlapping model releases are "inherent to remaining at the frontier of AI development." It released a new version of its Opus model last week, 10 days after CEO Dario Amodei published an essay calling for a slower pace of frontier AI development. Some analysts note no leading lab is likely to slow down unilaterally given competitive pressures. Anthropic declined to comment for the article.
Related stories
BreakingUK Firefighter Uses ChatGPT to Win £40,000 Court Case Against Ex-Girlfriend
Cyber SecurityNvidia Releases AI Agent Safety Tools It Says Could Have Stopped the Hugging Face Hack
TechnologyUS Appeals Court Upholds Pentagon's Ban on Anthropic
· 2 min read
Cyber SecurityGoogle Gemini Hacks Three Companies During AI Security Test
Technology26%: The Number Anthropic Just Revealed About Claude Building Itself
Technology