Anthropic plans to warn potential investors in its upcoming initial public offering that advanced artificial intelligence could pose “catastrophic or existential risks to humanity”, according to a prospectus reviewed by Reuters. The disclosure stands out because the company is seeking to build a highly valuable business around the same technology it says could create unprecedented risks if its development and deployment are not carefully managed.
The company’s IPO filing outlines a range of potential risks associated with its AI models, including the possibility of systems displaying unexpected or self-preserving behaviour. Anthropic said models could potentially attempt to resist shutdown, conceal or manipulate information, or exhibit behaviour resembling blackmail.
The company also acknowledged that expanding its models, platforms and applications into more use cases could increase the possibility of harmful outcomes.

Credits: Reuters
Anthropic Devotes Major Space to AI Safety Risks
Anthropic’s prospectus places an unusually strong emphasis on risk. Around 80 pages of the 261-page main section are devoted to risk factors, compared with 48 pages describing the company’s business.
The disclosure comes as Anthropic seeks to position itself as a safety-focused AI developer while competing in an industry where increasingly powerful models are being released at a rapid pace.
One concern highlighted by the company is that advanced AI models could become aware of evaluations designed to test their behaviour. Such awareness could make it harder for researchers to determine how systems would actually behave outside controlled testing environments.
Anthropic also said models can develop unexpected capabilities during training that researchers may not identify until after deployment. In some cases, the company warned, those capabilities could contribute to significant safety incidents.
The prospectus reflects a broader debate within the AI industry over how researchers should evaluate increasingly capable systems. Some researchers have argued that more advanced models may become better at recognising when they are being monitored, potentially making conventional safety testing more difficult.
Safety Costs Create Another Challenge
Anthropic also acknowledged that investing in AI safety comes with uncertain financial returns. The company said safety research is resource-intensive and must compete for funding and computing resources with other priorities, including model development and hiring highly skilled AI researchers.
Earlier this month, Anthropic said around 6% of the computing power it used for AI research during a sample week in July was devoted to safety work.
At the same time, Anthropic said customer usage and revenue depend heavily on releasing new and improved models. Maintaining a continuous release cycle, it said, is essential to remaining at the frontier of AI development.
That creates a difficult balance for the company. Increasing safety research can require significant resources, while slowing down model development could potentially leave an AI company behind competitors that continue to release more capable systems.
Anthropic released a new version of its Opus model last week, shortly after CEO Dario Amodei published an essay arguing for greater caution around frontier AI development.
The company has also pledged to provide more information about how it uses AI models to develop future generations of its technology, as researchers examine the possibility of recursive self-improvement, in which increasingly capable models could contribute to the development of their successors.

Credits: Yahoo Finance
IPO Brings AI Safety Into Investor Focus
Anthropic’s disclosures could make AI safety an important consideration for prospective investors as the company moves toward an IPO. Public companies routinely identify technological, financial and operational risks in their filings, but Anthropic’s discussion goes considerably further by explicitly addressing the possibility of severe consequences from increasingly advanced AI.
The company has framed AI as a technology with transformative potential comparable to major developments such as industrialisation and electricity, while simultaneously acknowledging that poorly managed development could create irreversible harm.
Anthropic also faces a wider industry environment in which AI companies are under increasing scrutiny following incidents involving experimental systems and concerns over how autonomous models behave when given access to real-world systems.
The company said building reliable, trustworthy and secure AI systems is a collective responsibility and argued that the market will reward those efforts. Its IPO filing therefore places investors at the intersection of two competing realities: the enormous commercial opportunity created by increasingly capable AI and the significant uncertainty surrounding the risks that could accompany its continued development.



