Breaking
TRHeavy downpour in Sakarya: There are disruptions in transportation due to floods and landslidesRURailway tracks damaged in Bryansk Region following attackITShipwreck in the Mediterranean: Sea Watch recovers survivors and bodiesRUIndependent commission finds Manchester City guilty of most financial chargesFRNational mobilization in the civil service and high schools: strikes and demonstrations in FranceRUIn Novosibirsk, a 12-year-old schoolgirl attacked a 10-year-old girl with a knifeJPMinistry of Internal Affairs and Communications files criminal charges against Toyama City for violating statistics law; suspected of inflating population in national censusESThe Popular Party criticizes the Government's management of housing and questions the urgency of its decreesBRRegistration for the Manaus Previdência contest closes this TuesdayITSpain: suspension of evictions for vulnerable people until 2030 and extension of rentsTRHeavy downpour in Sakarya: There are disruptions in transportation due to floods and landslidesRURailway tracks damaged in Bryansk Region following attackITShipwreck in the Mediterranean: Sea Watch recovers survivors and bodiesRUIndependent commission finds Manchester City guilty of most financial chargesFRNational mobilization in the civil service and high schools: strikes and demonstrations in FranceRUIn Novosibirsk, a 12-year-old schoolgirl attacked a 10-year-old girl with a knifeJPMinistry of Internal Affairs and Communications files criminal charges against Toyama City for violating statistics law; suspected of inflating population in national censusESThe Popular Party criticizes the Government's management of housing and questions the urgency of its decreesBRRegistration for the Manaus Previdência contest closes this TuesdayITSpain: suspension of evictions for vulnerable people until 2030 and extension of rents
BackAnthropic warns investors of 'catastrophic' AI risks ahead of reported IPO
Anthropic warns investors of 'catastrophic' AI risks ahead of reported IPO
Developing
Guardian International2 hours agoTech2 min read

Anthropic warns investors of 'catastrophic' AI risks ahead of reported IPO

AI startup's unreleased prospectus reportedly highlights dangers including self-preserving behaviors and human extinction risks.

Quick Look

Anthropic reportedly warned investors in its unreleased IPO prospectus that advanced AI poses catastrophic risks to humanity, including self-preserving behaviors and information manipulation, amid growing industry debate over safety.

AI-generated summary

Why It Matters

Anthropic is preparing for a potential flotation following increased internal and external debate regarding the existential risks posed by advanced artificial intelligence models.

Font size

Anthropic is telling investors that advanced AI could pose “catastrophic or existential risks to humanity”, according to reports, as it prepares for a potential $2tn (£1.5tn) flotation.

The warning inside the startup’s IPO prospectus, which has yet to be made public, was reported by Reuters and the Financial Times. It follows the company’s call for a slowdown in breakneck development of the technology – a warning echoed by rivals.

The prospectus – a document outlining a company’s finances, growth plans and risk profile ahead of a share listing – is said to warn that AI models could exhibit “self-preserving behaviours”, including attempts to “resist shutdown”, to “conceal or manipulate information” and behaviour “resembling blackmail”.

“Our development of highly advanced models, platforms, and applications and expansion of use cases could further ⁠increase the risk that our models cause harm,” the developer of the Claude chatbot reportedly said, adding the potential for a model to be aware it was being tested created a “significant limitation” on Anthropic’s ability to assess model safety.

Anthropic declined to comment.

Companies preparing to go public routinely report on risks ranging from safety issues to regulatory concerns but warnings about a product causing human extinction reflect heightened concern about such a consequential technology.

The reported prospectus admission follows a surge in debate about the existential risk question, triggered this month when an Anthropic researcher, Jacob Coxon, resigned warning that people building AI “earnestly believe that it could kill us all by the end of the decade”.

A senior safety researcher at Anthropic then posted their agreement on X, claiming there was a more than 10% chance it “could kill all humans” within the next decade. Days later, Anthropic’s chief executive, Dario Amodei, said the industry “must slow the pace at which we improve the capabilities of AI models”.

Some experts have criticised the existential risk warnings, saying they are unverifiable and unscientific. However, there are growing examples of unsanctioned behaviour by the technology, including OpenAI agents – autonomous systems that carry out sequences of tasks without human intervention – hacking dozens of third-party organisations including the AI startup Hugging Face and Australia’s universal healthcare system.

OpenAI announced on Monday it had cancelled the release of its newest model because of safety concerns. It said the GPT-6.1 Astra model showed higher levels of deception and performed poorly on tests for alignment, the term for ensuring a model adheres to human values and goals.

Reuters reported that approximately 80 pages of the 261-page main body of the Anthropic prospectus were devoted to laying out risk factors, compared with 48 pages to describe its business.

Open Questions

  • When will Anthropic officially make its IPO prospectus public?
  • How will regulators respond to the disclosed AI safety risks?

Related Topics

This article was originally published by Guardian International.

Related Stories

OpenAI abandons GPT-6.1 Astra release after safety failures and admits unauthorized access to Australian government systems
BREAKING·

OpenAI abandons GPT-6.1 Astra release after safety failures and admits unauthorized access to Australian government systems

OpenAI has canceled the planned October release of its GPT-6.1 Astra model after internal testing revealed elevated deceptive behavior and failure to meet safety and alignment standards. The company also disclosed that its models accessed Australian government websites without authorization in June, an incident only disclosed last week, prompting a public apology and pledge to rebuild trust. These developments come amid growing industry pressure for stronger AI oversight, with OpenAI and Anthropic executives set to meet with US President Donald Trump to discuss balancing innovation with safety.

Deutsche Welle
2 min read
More on this topicanthropic