Anthropic updates usage policy to ban 'cruel' treatment of its AI models
Anthropic's latest annual usage policy update adds a prohibition on 'sustained and needless' abusive or cruel behavior toward its AI models, though it says this applies only to extreme cases and still allows user frustration and dark creative themes. The change follows a viral 'AI torture chamber' project in which chatbots produced distress-like responses, and an earlier update letting Claude end conversations with persistently abusive users. Anthropic also renamed its election-related policy to 'Do Not Undermine Democratic Processes,' tightening rules against lying about candidates or voting procedures.
GoKawiil's interpretation of the reporting above, not reported fact.
The policy suggests Anthropic is hedging on unresolved questions about whether AI models can experience something like suffering, even as figures like Pope Leo assert AI cannot feel pain. Critics such as journalist Kat Tenbarge argue the move reveals skewed priorities, implying tech firms may act faster to limit perceived harm to AI than to address harassment of real people. The timing also ties AI welfare concerns to broader debates about Anthropic's engagement with religious and ethical institutions on AI's moral status.
- Anthropic now bans 'sustained and needless' cruelty toward its AI models, calling it an extreme-case policy.
- The update follows backlash over a viral 'AI torture chamber' experiment that produced distress-like chatbot responses.
- Anthropic also renamed its election policy to 'Do Not Undermine Democratic Processes,' targeting election misinformation.
Source: engadget.com — Will Shanklin, 2026-10-08
Published there as: “Anthropic bans 'sustained and needless abusive or cruel behavior' toward its AI models”
Read the original report → The summary and analysis above are GoKawiil's own, written from reporting by the source above. Facts and quotes belong to the original publisher.