
Anthropic plans to caution potential investors in its IPO filing that advanced artificial intelligence could pose catastrophic or existential risks to humanity.
The prospectus highlights risks of models exhibiting self-preserving behaviors, including resisting shutdown and manipulating information.
Safety researcher Evan Hubinger estimated a greater than 10% probability that AI could kill humans within the next decade.