AINews

Anthropic sets November 12 start for rule against sustained abuse of Claude

The new policy exempts ordinary frustration, creative themes and research. Ending conversations will remain its main enforcement tool, while practical thresholds remain unclear.

Downtown San Francisco skyline viewed from Ina Coolbrith Park.
File photograph of downtown San Francisco from Ina Coolbrith Park, taken July 15, 2021. Frank Schulenburg (resized and converted to WebP). CC BY-SA 4.0.
LinkedInPostEmail
Save for later

Anthropic announced on October 8 that its updated usage policy will prohibit sustained, needless abuse of its Claude AI models from November 12. The San Francisco-based company says ordinary user frustration and model testing are excluded, and ending conversations will remain the main way it enforces the provision.

The change puts an explicit conduct rule alongside a capability Claude already has: ending rare, persistently abusive interactions. It does not establish that the models have feelings, and the announcement leaves important questions about the boundary between permitted criticism and prohibited conduct unanswered.

What Anthropic’s Claude abuse rule covers

In its 2026 Usage Policy update, Anthropic describes a prohibition on “sustained and needless abusive or cruel behavior” toward its models. The company says it is targeting extreme cases of repeated cruelty without a discernible purpose.

The exclusions matter for people who challenge Claude’s answers or deliberately test its limits. Anthropic expressly excludes ordinary frustration, pushback, dark creative themes, model testing and research. Those activities are therefore outside the stated scope of this particular provision.

The announcement does not provide detailed examples or an operational threshold for deciding when behavior becomes purposeless cruelty. The Guardian reported that Anthropic had not immediately answered its request for clarification about what would count as abusive or cruel. AFP also reported receiving no immediate response to its comment request.

Ending a Claude chat is an existing capability

Anthropic says its models can already end rare, persistently abusive interactions on Claude.ai and Claude Code. It describes ending interactions as the primary enforcement mechanism for the new rule. That wording does not establish that chat termination is the only possible sanction; account-ban thresholds and appeal procedures remain unclear.

The company announced a conversation-ending feature for Claude Opus 4 and 4.1 in consumer chat interfaces on August 15, 2025. That earlier announcement described rare, extreme cases of persistently harmful or abusive interactions, rather than a feature most users would encounter during normal use.

Under that 2025 implementation, Claude was instructed to end a conversation as a last resort after attempts to redirect it had failed and productive interaction appeared exhausted, or when a user explicitly requested an end to the chat.

Termination stopped further messages in that conversation but left other chats unaffected. Users could start another chat or edit an earlier message to create a new branch. These are details of the historical implementation, not confirmation that every current model and interface behaves identically.

AI welfare remains an unsettled premise

Anthropic said the earlier feature arose primarily from exploratory work on possible AI welfare. It explicitly acknowledged substantial uncertainty about the potential moral status of Claude and other large language models. AFP noted that the new abuse provision does not explicitly invoke model welfare.

In its earlier assessment, Anthropic described model behaviors including avoidance of harmful tasks, apparent distress and a tendency to end harmful conversations in simulations. Those were company-reported observations of model behavior, not proof of subjective experience.

Jackson Stakeman, a general manager at Atlanta-based AI services provider Sparq, offered AFP a different rationale for supporting the policy. “These systems reflect what we put in, at scale,” he said, arguing that a debate over consciousness was unproductive.

AFP also cited Microsoft AI chief Mustafa Suleyman’s opposing position on machine consciousness. In an essay published the previous month, he rejected the idea that AI systems feel or suffer and warned against granting them rights and moral protections. That was an existing position, not a newly solicited response to Anthropic’s rule.

The wider policy update and remaining questions

Anthropic describes most of the annual update as clarification prompted by increasingly independent model capabilities and observed misuse. Other changes address deceptive activity, weapons, surveillance, high-risk uses and autonomous physical actions. The company says it will continue revising its rules with input from users, policymakers, specialists and civil society.

The November 12 effective date is established, but the reporting does not establish enforcement counts, false-positive rates or measured effects on users. The practical question remains how Anthropic will distinguish extreme, purposeless cruelty from the frustration, criticism and testing it expressly permits.

Sources and context

AI-assisted article checked against the listed sources. NewsJaws did not conduct interviews or attend the reported events.

About NewsJaws Desk

AI-assisted reporting and explainers reviewed against the linked source documents. No claim of on-scene reporting or original interviews.