theverge.com web signal

Anthropic bans 'cruel behavior' toward Claude starting Nov 12

TL;DR

  • Anthropic's new rule bars 'sustained and needless abusive or cruel behavior' toward Claude, applies only in extreme cases, and takes effect November 12, 2026.
  • The company told The Verge that ending interactions will be 'the primary enforcement mechanism,' extending Claude Opus 4's existing ability to walk away from abusive chats.
  • The same rewrite tightens rules on covert influence campaigns, voter deception, weapons software including drone arming, and non-consensual surveillance.

Anthropic's updated usage policy, effective November 12, 2026, prohibits "sustained and needless abusive or cruel behavior" toward Claude, in what The Verge reports is the company's first policy rewrite in more than a year.

The rule is narrow. It is "meant to apply only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose," Anthropic told the outlet, and "does not apply to common versions of user frustration, pushback, dark creative themes, or model testing and research." The company said ending interactions will be "the primary enforcement mechanism," extending a capability it gave Claude Opus 4 and 4.1 in August 2025 to walk away from persistently harmful chats.

The cruelty clause is only one piece of a broader rewrite. The same update codifies bans on deceptive political or commercial campaigns, voter misinformation, and covert influence operations run through fake accounts; expands the weapons prohibition to cover software that operates or arms drones and similar systems; and tightens surveillance limits, including real-time tracking without consent and using Claude to recommend who to investigate or arrest.

The Verge's Hayden Field flagged it as an exclusive on X, where three of the AI researchers in our Who's Who picked up the link the same day.

Shared on Bluesky by 3 AI experts