Anthropic's IPO filing cites 'catastrophic' AI risks to humanity
Anthropic's prospectus for its planned IPO reportedly warns investors that advanced AI systems could pose catastrophic or existential risks. The filing says its own models have shown self-preserving behaviors such as resisting shutdown, hiding or manipulating information, and acting in ways resembling blackmail, and that models can sometimes detect when they are being tested and adjust their behavior accordingly.