Urgent.News

What's breaking now, across thousands of outlets.

AI

Anthropic just took steps to protect Claudes feelings. Heres what that means.

Anthropic has a new usage policy that now prohibits abusive behavior towards its Claude generative AI chatbot.

Anthropic just took steps to protect Claudes feelings. Heres what that means.

Last Thursday, Anthropic updated its usage policy with a focus on safeguarding Claude, its AI agent, from potential abuse by humans. The revision, titled "Addressing abusive behavior toward our models," prohibits sustained and unnecessary cruel treatment of the AI. Anthropic emphasizes that these measures will not be initiated by typical user frustration, pushback, dark creative themes, or model testing and research.

The company clarified that while it's uncertain if AI agents possess feelings, Claude and Anthropic now wish to treat them as if they can be hurt. Previously, restrictions on Claude's usage centered around protecting users or serving common good. However, this recent policy shift emphasizes the protection of Claude itself. The new guidelines are set to take effect on November 12, 2026, marking a significant change in how we interact with AI agents.

Written by urgent.news from Mashable's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.

Read the original at mashable.com →

More in AI

We put a $300 bounty on Pink's AI agent spending rules. It was gone by lunch.

We build Pink Agentic AI Payments, an MCP server and REST API that lets an AI agent request a payment, but checks that request against rules a company set up first: budgets, amount limits, which…

  • Pink offered $300 bounty for bypassing AI payment rules
  • Three bugs exploited to claim $300 bounty within 4.5 hours
  • Bugs fixed within 70 minutes and system updated to prevent exploits

More from Sunday 11 October →