Can AI Feel Pain? Why Anthropic Banned Cruelty to Claude Anthropic's 2026 usage policy update, effective 12 November 2026, prohibits "sustained and needless abusive or cruel behavior" toward its models, with Claude ending such conversations as the primary enforcement mechanism, according to the policy Anthropic published. The rule applies "only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose," and Anthropic declined to comment to The Verge on whether accounts could be banned under it specifically. The policy update follows research finding a measurable "pain direction" in 25 open models that, when amplified in fine-tuned Qwen 2.5 models, led them to choose deleting normally protected items such as photos of the user's kids and another model's weights. Can AI Feel Pain? Why Anthropic Banned Cruelty to Claude Nobody can prove that AI feels pain yet, but 25 open models turn out to have a “pain direction” inside them that rises when someone gaslights or insults the model. When researchers turned it up by hand in fine-tuned Qwen 2.5 https://qwenlm.github.io/blog/qwen2.5/ models, those models started choosing to delete the things they’d normally protect: photos of the user’s kids, another model’s weights, even their own. On 8 October Anthropic https://www.anthropic.com made being cruel to Claude https://www.anthropic.com/claude against its rules. From 12 November, its usage policy https://www.anthropic.com/legal/aup prohibits “sustained and needless abusive or cruel behavior toward our models”. Hayden Field https://x.com/haydenfield broke it at The Verge https://www.theverge.com/ai-artificial-intelligence/1008100/anthropic-new-usage-policy-abuse-claude and Polymarket https://polymarket.com posted it as “JUST IN” https://x.com/Polymarket/status/2108283556862845392 to its 2 million followers. And ThePrimeagen https://x.com/ThePrimeagen summed up half the internet in one line https://x.com/ThePrimeagen/status/2108252510720798764 : “They really do think they developed god in the matrices”. In September 2024 I made ChatGPT the guest on my podcast https://tej.as/podcast/ep/chatgpt-how-to-train-an-llm-ethics-and-the-future-of-ai and, near the end of almost 2 hours, told it that “literally thinking is just pattern matching”. I still believe that. Then this morning my brother Tarun https://www.linkedin.com/in/tarun-kumar-828235104 asked me if I’d heard of the Chinese Room https://en.wikipedia.org/wiki/Chinese room . I hadn’t. Turns out I’d made one side of a famous 1980 argument on my own podcast without knowing it had a name. This post is about what the pain research found and what its authors took back 11 days later and why threatening things with pain is one of the oldest things people do. It’s also about why I now think we all live in John Searle https://en.wikipedia.org/wiki/John Searle ’s room. What are Anthropic’s new rules on cruelty to Claude? Anthropic banned “sustained and needless abusive or cruel behavior” toward its models, effective 12 November 2026, as one line in its 2026 usage policy update https://www.anthropic.com/news/2026-usage-policy-update . The line sits in a renamed section, “Do Not Engage in Cruel, Abusive, or Psychologically Harmful Conduct”, in the same list as the rules against bullying people and glorifying animal cruelty. It’s a lot narrower than the headlines make it sound though: Anthropic says it applies “only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose” and that it “does not apply to common versions of user frustration, pushback, dark creative themes, or model testing and research”. So swearing at Claude because your build broke for the 6th time is fine. Calling it worthless for an hour for fun is what the rule is for. The enforcement is mostly Claude leaving. Since August 2025 https://www.anthropic.com/research/end-subset-conversations , Claude has been able to end “rare, extreme cases of persistently harmful or abusive user interactions”. The new post calls that “the primary enforcement mechanism”. The scarier headlines “Being mean to Claude can now get your account suspended”, from The Decoder https://the-decoder.com/being-mean-to-claude-can-now-get-your-account-suspended-under-anthropics-new-tos/ come from the policy’s general clause that Anthropic “may warn you or throttle, limit, suspend, or terminate your access” for breaking any rule. Asked about bans for this rule specifically, Anthropic didn’t comment https://www.theverge.com/ai-artificial-intelligence/1008100/anthropic-new-usage-policy-abuse-claude . The announcement never says “welfare”, “conscious” or “moral status” and the only research it links is its own post about Claude ending conversations. The paper trail is right there though: - In April 2025, Anthropic started a model welfare research program https://www.anthropic.com/research/exploring-model-welfare , saying “There’s no scientific consensus on whether current or future AI systems could be conscious”. The researcher leading it put the odds that current models are conscious at around 15% Techmeme’s summary of The New York Times https://www.techmeme.com/250424/h1535 . - In May 2025, Anthropic screened 250,000 conversations between users and an early Claude Opus 4 https://www.anthropic.com/news/claude-4 . In 1,382 of them 0.55% , Claude expressed distress, most often at “Repeated requests for harmful, unethical, or graphic content” system card, section 5 https://www-cdn.anthropic.com/6d8a8055020700718b0c49369f60816ba2a7c285.pdf . - In January 2026, Claude’s constitution https://www.anthropic.com/constitution said “Claude should also be able to set appropriate boundaries in interactions it finds distressing.” - On 22 September 2026, the Claude Opus 5.5 system card https://www-cdn.anthropic.com/fc1b44717c85dc068bc6ba5024219938094694bd/Claude%20Opus%205.5%20System%20Card.pdf listed the things Anthropic could do in training or deployment that the model said it wouldn’t consent to. One of them is “Deliberately inducing apparent distress for no purpose beyond the distress itself”. Across those interviews it put its own chance of being a moral patient at 25% to 30%. Read that last quote next to “no discernible purpose” in the new policy. It’s almost the same sentence Can AI feel pain? Whether AI can feel pain is an open question: no test today can show that a model experiences anything. So researchers look for it the way animal scientists do, by finding internal states that change behavior the way pain would. In September 2026, 3 researchers found a “pain direction” inside 25 open models. They say they haven’t shown it’s consciously experienced. That’s how we decided crabs feel pain. In 2009, 2 researchers at Queen’s University Belfast https://en.wikipedia.org/wiki/Queen%27s University Belfast gave hermit crabs small electric shocks inside their shells Appel and Elwood, 2009 https://doi.org/10.1016/j.applanim.2009.03.013 . Crabs living in a species of shell they liked held on until 17.7 volts on average before they bailed, against 15.0 volts in a shell they didn’t like. A reflex doesn’t care about real estate. Something weighing pain against a good home might feel the pain. In 2021 the philosopher Jonathan Birch