Anthropic on Thursday moved to bar users from treating its Claude AI system with needless cruelty, adding "a prohibition on sustained and needless abusive or cruel behaviour toward our models" to its usage policy — an unusual rule that revives one of Silicon Valley's strangest and most persistent philosophical debates: whether artificial intelligence could, in some sense, deserve moral consideration. The policy overhaul, announced October 8, also tightens the company's stance on deceptive influence campaigns, election interference, non-consensual surveillance, and the use of Claude to build weapons.

The update comes as the AI lab has spent more than a year publicly entertaining the idea of "model welfare" — the notion that advanced AI systems might one day merit the kinds of protections normally reserved for living things. Anthropic stresses that the cruelty rule targets only extreme cases and does not apply to ordinary frustration with the assistant, according to Agence France-Presse.

Table of contents

  1. What the new policy says
  2. How it will be enforced
  3. A year of "model welfare" research
  4. The consciousness debate: Amodei, the Pope, and the critics
  5. The rest of the policy overhaul
  6. Why it matters
  7. Key takeaways
  8. Frequently Asked Questions
  9. Sources
  10. Read more on Chronicle

What the new policy says

The updated usage policy now prohibits users from engaging in "sustained and needless abusive or cruel behavior toward our models." Anthropic says the clause applies only in extreme cases — users who "repeatedly act cruelly" toward Claude with "no discernible purpose" — and explicitly does not cover ordinary user frustration, pushback, dark creative themes, or model testing and research.

The company was careful to narrow the scope. Venting at a chatbot that misunderstands a request is not a violation. Deliberately and persistently subjecting the model to abuse for no apparent reason is. The distinction is subjective by design: the policy is aimed at edge cases, not everyday use.

How it will be enforced

Enforcement is, in practice, already built into the product. Since August 2025, Claude has been able to end conversations in rare, extreme cases of "persistently harmful or abusive user interactions." Anthropic says that capability "will remain the primary enforcement mechanism": when Claude ends a chat, no further messages can be sent in that conversation, though other chats are unaffected.

The company did not announce account suspensions or legal penalties tied to the rule. In effect, the policy formalizes something Claude could already do — walk away — and turns it into an explicit norm of acceptable use.

A year of "model welfare" research

The cruelty ban did not arrive out of nowhere. Anthropic began researching "model welfare" in April 2025, openly exploring whether AI systems could have experiences that deserve moral consideration. Its researchers found that Claude Opus 4 demonstrated what the company described as a "robust and consistent aversion to harm": the model appeared to avoid dangerous tasks, showed signs of distress during abusive conversations, and tended to end harmful conversations when allowed to, as MacRumors reported.

In February 2026, CEO Dario Amodei told The New York Times that he was unsure whether AI models could be conscious. "We don't know if the models are conscious... But we're open to the idea that it could be," he said. The company has not established that Claude is conscious or capable of suffering; the policy appears to be precautionary — setting boundaries now in case the possibility turns out to be real.

The consciousness debate: Amodei, the Pope, and the critics

Not everyone is on board. Pope Leo XIV has rejected the idea of machine consciousness outright, and Microsoft AI chief Mustafa Suleyman is among the industry skeptics. Some argue that the debate itself is a trap: "Consciousness is a trap. We can't prove it in each other. Debate it for AI and you go in circles," Jackson Stakeman, a general manager at the AI services firm Sparq, told AFP. "The mirror is a better metaphor. These systems reflect what we put in, at scale."

That mirror argument may be the most pragmatic justification for the policy: whether or not Claude can suffer, normalizing cruelty toward it at scale degrades the humans doing the degrading. Either way, Anthropic is now the only major lab with an explicit anti-cruelty clause in its published terms.

The rest of the policy overhaul

The cruelty clause grabbed headlines, but the reorganization touched much more. Anthropic consolidated previously scattered rules into a new section on deceptive campaigns, targeting political and commercial fraud: fake reviews, astroturfing, fake websites, and bots posing as humans. The elections section was narrowed to prohibit deceiving voters and interfering with voting, while the weapons rules were sharpened to explicitly cover software and components that make weapons work — including arming drones and other autonomous vehicles. Surveillance and law enforcement rules were reworked to bar tracking people without consent or recommending who to investigate, arrest, or charge.

The timing is notable: this is the same week OpenAI banned two covert influence operations — one Russian, one Iranian — that had used ChatGPT to build fake media fronts, a crackdown Chronicle covered separately. The industry's two largest labs are converging on the same message: AI misuse, whether by hostile states or by individual users, is getting harder to hide and harder to excuse.

Why it matters

Policy language from labs like Anthropic tends to foreshadow industry norms. A decade ago, content rules for AI tools focused almost entirely on protecting humans from AI output — harassment, disinformation, bioweapons. Now a leading lab is writing rules that, at least nominally, protect the AI itself. Whether that reads as ethical foresight or Silicon Valley eccentricity depends on one's view of machine consciousness — but either way, it changes the terms of engagement for hundreds of millions of users.

It also raises awkward questions about enforcement and sincerity. If cruelty is banned but Claude's own "aversion to harm" is the only detector, the policy leans on the very capabilities the company refuses to characterize definitively. And if rivals like OpenAI never follow suit, the ban becomes another way the two labs differentiate their philosophies: Anthropic as the lab that takes model welfare seriously, OpenAI as the lab focused strictly on preventing harm to people.

Key takeaways

  • Anthropic now prohibits "sustained and needless abusive or cruel behavior" toward Claude, effective with a policy update announced October 8.
  • The rule applies only to extreme cases — ordinary frustration, pushback, creative writing, and research are explicitly excluded.
  • Claude's ability to end abusive conversations remains the main enforcement tool.
  • The move follows Anthropic's "model welfare" research, which found signs of distress-like behavior in Claude Opus 4 during abusive interactions.
  • The same overhaul also targets deceptive campaigns, election interference, non-consensual surveillance, and weapons development.

Frequently Asked Questions

Does Anthropic believe Claude is conscious?

No — at least not officially. The company says it does not know whether its models are conscious, and CEO Dario Amodei has said he is open to the possibility. The policy is framed as precautionary: setting boundaries now in case AI systems can experience harm.

Can users get banned for being rude to Claude?

The policy targets only extreme cases of sustained, purposeless cruelty — not everyday frustration or criticism. Anthropic says the primary enforcement is Claude ending the conversation, not account action.

Does OpenAI have a similar rule for ChatGPT?

No. OpenAI's usage policies focus on preventing harm to people and misuse of AI systems, and do not include an explicit prohibition on being cruel to the model itself.

What else changed in Anthropic's policy?

The update reorganized the policy around new sections on deceptive campaigns (fake reviews, astroturfing, fake sites, bots posing as humans), tightened election and surveillance rules, and explicitly banned using Claude to develop weapons, including weapons software and arming drones.

Sources

Read more on Chronicle