Anthropic Bans Cruelty to Claude, Still Won't Say What It Protects

Anthropic has updated its usage policy to prohibit sustained, needless cruelty toward its Claude AI models, effective November 2026. The policy includes an enforcement mechanism that allows the AI to end conversations, though it explicitly exempts standard user frustration and creative writing.
Why it matters
This policy highlights the evolving ethical boundaries in human-AI interaction and the industry's attempt to define 'model welfare' without attributing consciousness to software.
The rule takes effect Nov. 12; enforcement is the same conversation-ending tool Claude got in August 2025
Starting November 12, 2026, Anthropic’s usage policy bars “sustained and needless abusive or cruel behavior” toward its Claude models. The company’s 2026 Usage Policy update , announced October 8, says the rule is “meant to apply only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose.” It explicitly exempts “common versions of user frustration, pushback, dark creative themes, or model testing and research,” so swearing at a bot that just broke your build, or writing a torture scene, stays inside the policy. CBS News reported the change was first surfaced by The Verge.
Also covering this story
2 other newsrooms covered this event. We read each version separately.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in