Starting November 12, 2026, Anthropic’s usage policy bars “sustained and needless abusive or cruel behavior” toward its Claude models. The company’s 2026 Usage Policy update, announced October 8, says the rule is “meant to apply only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose.” It explicitly exempts “common versions of user frustration, pushback, dark creative themes, or model testing and research,” so swearing at a bot that just broke your build, or writing a torture scene, stays inside the policy. CBS News reported the change was first surfaced by The Verge.
The enforcement mechanism isn’t new. Anthropic gave Claude Opus 4 and 4.1 the ability to end a conversation in August 2025, framed at the time as “part of our exploratory work on potential model welfare,” according to Anthropic’s own announcement. The company said this was a last resort, triggered only after multiple refusals and redirects failed, and that “the vast majority of users will never experience Claude ending a conversation.” There’s a hard override: Claude won’t end a chat if a user seems at risk of harming themselves or others. No figures on how often the tool actually fires have been published, in 2025 or in the new policy, so there’s no way to check the “vast majority” claim against a rate.




