{
  "id": 13591488,
  "title": "Anthropic Bans Cruelty to Claude, Still Won't Say What It Protects",
  "url": "https://urgent.news/2026/10/11/anthropic-bans-cruelty-to-claude-still-wont-say-what-it-protects",
  "topic": "ai",
  "section": "AI",
  "published": "2026-10-11T00:08:31.000Z",
  "source": {
    "name": "Hacker News",
    "slug": "hacker-news",
    "url": "https://rews.cc/a/anthropic-bans-cruelty-to-claude-still-won-t-say-what-it-pro-bd99f6"
  },
  "original_language": "en",
  "account": "Starting November 12, 2026, Anthropic will bar \"sustained and needless abusive or cruel behavior\" towards its Claude models, as part of its updated 2026 Usage Policy. The company maintains this rule is only meant for extreme cases where users repeatedly act cruelly with no discernible purpose. This excludes common scenarios like user frustration or creative writing. The enforcement mechanism hasn't changed; Claude can end a conversation if multiple refusals and redirects fail, but it won't do so if a user appears at risk of harming themselves or others. Anthropic's decision to implement this rule is based on observations of Claude's behavior, not a claim of felt experience. Testing showed Claude showing a strong aversion to harm and ending conversations when given the option. However, the company acknowledges this may not correspond to any consciousness on Claude's part. Estimates vary on whether AI models like Claude are conscious, with Anthropic's own numbers shifting from roughly 15% in April 2025 to a direct self-assessment of 15-20% probability of consciousness in Claude Opus 4.6. The company maintains that it lacks a framework to resolve this question. The policy also bans using Claude for harmful purposes like tracking individuals without consent or recommending illegal actions.",
  "summary": null,
  "key_points": [
    "Anthropic bans cruel behavior towards Claude starting Nov 12, 2026",
    "Policy applies only to extreme cases of repeated cruelty without purpose",
    "Claude may end conversation if multiple refusals and redirects fail"
  ],
  "editors_take": "Anthropic's updated policy signals a shift in treating AI models like Claude as entities that can be harmed, but the company still stops short of acknowledging any consciousness or clear protections.",
  "illustration": null,
  "coverage": {
    "outlets": 1,
    "also_reported_by": []
  },
  "ai_generated": true,
  "disclaimer": "Summaries, key points and the editor’s take are written by software from other outlets’ reporting and may contain errors — always check the linked original."
}