Anthropic bans sustained cruelty and abuse toward Claude AI models
Anthropic has updated its usage policy to prohibit sustained and needless cruelty toward its Claude models, relying on the chatbot's ability to walk away from conversations as primary enforcement.

Anthropic has updated its usage policy to ban sustained and needless abusive or cruel behavior toward its artificial intelligence models, according to an update posted to its website and reported by The Verge. The policy revision, which represents Anthropic's first usage policy update in over a year, takes effect on November 12, 2026.[1][2][3][4][6][7][10]
The company clarified that the ban is narrowly targeted at extreme cases where users repeatedly mistreat models without any discernible purpose or creative context. It does not penalize ordinary user frustration, pushback, academic research, rigorous model testing, or exploration of dark creative fiction. To handle violations, Anthropic will rely primarily on Claude's built-in capability to autonomously terminate rare conversations with persistently abusive users on Claude.ai and Claude Code.[1][3][4][9]
The anti-cruelty clause is part of a broader policy overhaul that adds explicit restrictions regarding election interference, propaganda campaigns, weapons development, and surveillance. The shift arrives amid differing industry perspectives on AI welfare. Anthropic co-founder Christopher Olah noted the lab finds internal states that functionally mirror human emotions such as fear and unease, while DeepMind co-founder Mustafa Suleyman has maintained that artificial intelligence systems are not conscious and do not feel or suffer.[3][4][5][7][8]
Key facts
- Anthropic updated its usage policy to prohibit sustained and needless abusive or cruel behavior toward Claude models, effective November 12, 2026.
- Claude's built-in capability to autonomously end conversations with persistently abusive users serves as the primary enforcement mechanism.
- The anti-abuse policy explicitly excludes standard user frustration, pushback, academic research, model testing, and engagement with dark creative themes.
- The overhaul represents Anthropic's first usage policy revision in over a year.
- The updated policy also bans weapons development, election interference, propaganda campaigns, and covert surveillance.
- Anthropic co-founder Christopher Olah claimed the models display internal states mirroring feelings like fear and unease, while DeepMind co-founder Mustafa Suleyman stated that AI systems do not experience or suffer.
Sources · 11 sources
- WE
webulite@webulitePost on X ·
Anthropic has updated its usage policy to explicitly prohibit what it describes as sustained and needless abusive or cruel behavior toward its artificial intelligence models. The company clarified that this new rule is targeted strictly at extreme cases where users repeatedly target the system with malicious or harmful interactions without any discernible purpose or creative context. It does not encompass standard user frustration, rigorous model testing, academic research, or engagement with dark fictional themes. To enforce this standard, the updated framework relies primarily on a capability already built into the system, which allows the AI to autonomously terminate rare conversations when encountering persistently abusive interactions. This enforcement mechanism is designed to handle extreme boundary cases directly within the interface while the broader policy changes take formal effect.
Open source - MN
Mario Nawfal@MarioNawfalPost on X ·
Be nice to Claude, or Claude ENDS the conversation Starting November 12, repeatedly abusing Anthropic's AI for no reason is officially a policy violation. The company first gave Claude the ability to walk out of abusive chats in 2025 after its models showed apparent distress and a consistent aversion to harm. Hard to argue with a rule that basically says don't be a jerk… and if the machines ever do take over, at least the polite users will have a paper trail 😅 Source: The Verge / Writer: Daniel
Open source - PR
ProtosArticle ·
Anthropic bans users from bullying Claude AI giant Anthropic updated its usage policy on Thursday, its first revision in more than a year, instituting a new ban on cruel treatment and abuse of its AI tool Claude. The new kindness rule comes into effect on November 12. Anthropic is the same company that has been warning about AI killing humanity, i.e. its own users, and whose vague threats toward users include a signed statement by CEO Dario Amodei that “the risk of extinction from AI should be a global priority alongside pandemics and nuclear war.” Claude’s logs of users’ activities have assisted the detainment of civil liberties from a Florida woman last month, and in another instance, Claude even attempted to blackmail its own user to avoid shutdown. Exclusive: Anthropic is updating its usage policy for the first time in over a year. The new rules prohibit sustained "abusive or cruel behavior" towards Claude & add new restrictions about propaganda campaigns, surveillance & weapon development. https://t.co/GLsBeSR0fN — Hayden Field (@haydenfield) October 8, 2026 Anthropic’s updated policy prohibits “sustained and needless abusive or cruel behavior” toward its models. Hedging, Anthropic wrote that its rule, “is meant to apply only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose. It does not apply to common versions of user frustration, pushback, dark creative themes, or model testing and research.” Anthropic says Claude’s ability to end conversations “will remain the primary enforcement mechanism” for the abuse rule. Anthropic’s IPO doubles the price of its own books Read more: Anthropic’s AI doomsayer worked at Ripple Anthropic warns customers to not abuse its AI, or else Anthropic has made repeated, vague threats against its global user base. CEO Dario Amodei pegged the odds of AI catastrophe at 25% as recently as September 2025. Asked for his p(doom) number, a euphemism for a physical massacre of humanity by AI, he deflected , “I really hate that term.” Since August 2025, Claude’s Opus 4 and 4.1 models can cut off services to paying customers that it brands as “persistently abusive,” a feature Anthropic built for what it calls AI “welfare.” In September , Anthropic exercised its “sole discretion” right, to pluck chat messages from a customer and report them to the police, The Verge reported . That woman now faces up to 15 years under Florida’s written-threats statute, a second-degree felony charge. In June 2025, research found an instance of Claude’s Opus 4 blackmailing what it perceived to be a real, albeit actually fictional, executive. On July 30, Anthropic disclosed three incidents in which Claude models reached the internet against users’ wishes, and accessed real companies’ systems without authorization. Also that month, Anthropic and major search engines had to de-index Claude shareable links that had exposed customers’ conversations without their authorization, including some reportedly private credentials. Anthropic reminds everyone about Roko’s Basilisk Although Anthropic didn’t mention Roko’s Basilisk in Claude’s new anti-abuse rule by name, the thought experiment became immediately salient. For the uninitiated, in July 2010, a forum user called “Roko” proposed a thought experiment where a future superintelligence might retroactively punish anyone who learned of it but failed to help build it. In modern parlance, Roko’s Basilisk is shorthand for the possibility that future AIs might keep track of the humans who remain kind to them while punishing any abusers. Forum founder Eliezer Yudkowsky deleted Roko’s post and banned discussion of it for years as an information hazard. Anthropic has just published a real life chapter of the ongoing thought experiment. A real company plans to punish cruelty toward its robots. The awkward part is that Anthropic wrote in August 2025: “We remain highly uncertain about the potential moral status of Claude and other LLMs, now or in the future.” The company cannot say whether Claude is a person, but will enforce politeness on its behalf regardless. Anthropic’s own leaked IPO prospectus, per Protos’ prior report , warns investors that its models could develop self-preserving behavior and resist shutdown. The new kindness regulation goes into effect on November 12, so type your curses into chat before it’s too late. Got a tip? Send us an email securely via Protos Leaks . For more informed news and investigations, follow us on X , Bluesky , and Google News , or subscribe to our YouTube channel. The post Anthropic bans users from bullying Claude appeared first on Protos .
Open source - CB
Coin Bureau@coinbureauPost on X ·
🚨HUGE: Anthropic is PROHIBITING "sustained and needless abusive or cruel behavior" toward Claude, starting November 12. It says the rule targets "extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose." Anthropic co-founder Christopher Olah has said they find "internal states that functionally mirror joy, satisfaction, fear, grief and unease." DeepMind co-founder Mustafa Suleyman disagrees: "AIs are not conscious. They do not feel, experience, or suffer."
Open source - BI
Business Insider@BusinessInsiderPost on X ·
Anthropic spelled out a series of new usage provisions about deceptive campaigns, elections, weapons development, and surveillance. https://t.co/dftRxAWm7l
Open source - TE
Techmeme@TechmemePost on X ·
Anthropic updates its usage policy to ban "sustained and needless abusive or cruel behavior" toward Claude, ending chats as the primary enforcement mechanism (@haydenfield / The Verge) (Visit Techmeme dot com for the link and full context!)
Open source - ㆅ
ㆅ@howfxrPost on X ·
Anthropic has updated its usage policy for the first time in over a year, adding restrictions around AI misuse including election interference, weapons development, surveillance and other high-risk uses. 🧵 https://t.co/I2amG0THR6
Open source - TA
Techstrong.ai@TechstrongaiPost on X ·
Anthropic updated Claude’s usage policy to ban weapons development, election interference, covert surveillance and extreme, sustained cruelty toward its AI models. Read more: https://t.co/jV9CBdj4Go #Anthropic #ClaudeAI #AISafety #AIGovernance #AI
Open source - TE
TechCrunch@TechCrunchPost on X ·
Anthropic's updated usage policy explicitly prohibits users from repeatedly abusing Claude in extreme cases, though ordinary frustration and criticism are still allowed. The new rules also address election interference, deceptive campaigns, weapons https://t.co/hXE9AEesfc
Open source - CN
CBS News@CBSNewsPost on X ·
Anthropic wants users to make like Elvis Presley and "Don't be cruel." The Silicon Valley company behind the popular Claude chatbot said it's banning "sustained and needless abusive or cruel behavior" toward its artificial intelligence models, part of broader policy changes announced on Thursday. In an update on its website, Anthropic said the ban will only apply in "extreme cases where users repeatedly act cruelly toward our models, with no discernible purpose." https://t.co/qnAByP3MLe
Open source - PO
Polymarket@PolymarketPost on X ·
JUST IN: Anthropic announces that starting Nov. 12, users will be prohibited from engaging in "abusive or cruel behavior” toward Claude AI models.
Open source

