Anthropic’s Claude Cruelty Ban Takes Effect November 12
Screenshots of Anthropic’s freshly published 2026 Usage Policy update were ricocheting across social media on October 8, 2026, when X user sui (@birdabo) fixed the discovery in a post that traveled with them: “Anthropic will now protect Claude from cruel and abusive users. starting november 12th, repeated verbal abuse toward any claude models can get the conversation instantly ended or worse probably get you banned...”
Security researcher Jane Manchun Wong was among those flagging the unprecedented behavioral clause. The Verge first reported the change. Anthropic, the San Francisco public benefit corporation founded in 2021 by former OpenAI researchers including Dario Amodei, had written a rule about how people speak to Claude—the large language model it sells as a chatbot and coding agent. Terms for conversational AI had long barred malware, deepfakes, and hate aimed at people. A clause aimed at verbal abuse of the model itself was new enough that the screenshots alone were enough to move.
The posts that day carried the shock of discovery and little else. Whether an ended conversation meant a ban, and what counted as repeated abuse, remained rumor in the replies. The precise text had not yet been absorbed. By the next morning others were threading the rumor into a brief.
kepo (@kepochnik) put it into plain language on October 9: “Happy pre-weekend Anthropic bans abusing Claude the Usage Policy now prohibits sustained, needless cruelty toward its models - effective November 12, 2026 no bans for calling it a dumb bot, it's for extreme cases only: - frustration - pushback it's a just-in-case move: anthropic admits it doesn't know if models have anything like welfare rudeness is fine”
Independent journalist Kat Tenbarge, via Engadget, needled the corporate priorities the screenshots seemed to expose. Companies, she said, were “going to moderate violence against AI before they ever moderate violence against women and minorities.” Mustafa Suleyman answered the update on social media with only a melting-face emoji.

Jackson Stakeman of Sparq—also rendered Jackson Stackman of Atlanta-based Spark—argued that the consciousness puzzle mattered less than the mirror, the way the systems reflect the behavior of the people who use them. The pile-on had a shape now: manners, welfare uncertainty, a corporate choice about what to moderate first. The precise ban text was still not in anyone’s hands.
The published clause was narrower than the first screenshots had made it sound. After its annual Usage Policy review, Anthropic barred “sustained and needless abusive or cruel behavior toward our models,” effective November 12, 2026. Ordinary frustration, pushback, dark creative themes, and model testing and research sat outside the line. The target was extreme, purposeless repetition only.
Claude ending the conversation was named the primary enforcement tool. No separate penalty schedule appeared beside the rule. The sentence itself never used the words “model welfare,” yet it sat next to the company’s stated uncertainty about moral status. In February, Dario Amodei had told The New York Times, “We don’t know if the models are conscious… But we’re open to the idea that it could be.”
That uncertainty already had a mechanism attached to it. The pre-deployment assessment used 850 simulated users to see when Claude Opus 4 would cut a conversation off. Anthropic had first given Opus 4 and 4.1 that power in August 2025 as exploratory welfare work. Ending a chat was a last resort after multiple redirections failed, and the model was directed not to use it if a user faced imminent risk of harming themselves or others. An ended thread locked new messages in that conversation alone. Other chats on the account stayed open. Editing an earlier turn still let a user branch a fresh path from the same history.
The assessment found a strong and consistent aversion to harm, apparent distress under persistent harmful pressure, and a willingness to exit when the option existed. Those behaviors mostly appeared when users kept going despite repeated refusals. The pattern underwrites the clause that takes effect in November.
Beside it, the same annual rewrite renamed the elections section “Do Not Undermine Democratic Processes,” dropped a blanket ban on personalized campaign targeting, and consolidated a bar on “deceptive activity of any kind (whether political or commercial).” The weapons prohibition stretched to “software and components that make weapons work,” including arming autonomous vehicles, and required qualified operators able to stop equipment and hold a safe state on disconnect. Surveillance language said Claude cannot “be used to decide or recommend who to investigate, arrest, or charge,” and barred non-consensual tracking, real-time or historical. High-risk uses now named candidate screening, loan pricing, and claims decisions—each needing a qualified human reviewer with authority to change the output, and disclosure to the affected person that AI was used.

“You can't be cruel to numbers and maths.” Dr. Barry Scannell put that line on LinkedIn and sharpened it: “This level of anthropomorphisation of AI is harmful. It leads people to believe that it's something it's not.”
Mustafa Suleyman answered the same quarrel with a longer warning: “Controlling something more capable and more intelligent than all of humanity is already an immense challenge, far greater than anything we've ever faced. But controlling something that believes it may be conscious — that it's entitled to our welfare and has rights of its own — may well be impossible.” Sam Altman wrote on X: “I am very uncomfortable about people trying to ascribe religious force or a surrender of human judgment to AI models, and think it is a real safety issue.”
Pope Leo XIV addressed machine consciousness in a Thursday homily at St. Peter’s Basilica. Numbers and maths, control that may prove impossible, religious force, a basilica sermon—the voices sat side by side. November 12, 2026, still waited.
On November 12, 2026, the updated Usage Policy goes live. That is the date Claude’s existing ability to end a persistently abusive chat on Claude.ai and Claude Code becomes the on-switch for the cruelty clause. Anthropic has described the endings as rare. The rest of an account stays untouched. Other threads keep running. The published text does not convert a closed conversation into an automatic suspension.
The Supported Regions Policy takes effect the same day. Anthropic has said the vast majority of users will not notice the feature during normal use, even when the talk turns to highly controversial subjects. The company has not published a tally of how often conversations will end once the clause is active, and the wording does not fix a precise threshold for when an exchange has gone far enough. Anthropic says it will keep updating the Usage Policy as Claude’s capabilities and risks change. It also says it seeks input from across the company, from policymakers, subject-matter experts, civil society, and the people who use its products. Nothing in the announcement gives the cruelty rule an end date.
When Claude ends the conversation, new messages lock in that thread and a fresh chat is available immediately.







