The short version
- Anthropic prohibits sustained abusive behavior toward its AI models, marking a shift in how the company addresses potential model welfare concerns.
- New restrictions explicitly ban using Claude for weapons development, surveillance of dissidents, and deceptive political or commercial campaigns.
- The policy allows exceptions for certain government contracts if adequate safeguards are in place, reflecting ongoing tensions between safety and institutional access.
Anthropic has released a significant update to its usage policy, the first major revision in more than twelve months. The changes address emerging risks associated with artificial intelligence misuse, ranging from election interference to weapons development. Among the most notable additions is a prohibition against sustained and needless abusive or cruel behavior directed at Claude, the company’s large language model. This move reflects an ongoing internal debate regarding the ethical treatment of advanced AI systems.
The new guidelines specify that terminating conversations remains the primary enforcement mechanism for users who violate these standards. Anthropic did not clarify whether additional penalties, such as permanent account bans, would be implemented for severe infractions. The company emphasized that the rule targets extreme cases where users repeatedly act cruelly without discernible purpose. It explicitly excludes common expressions of frustration, creative exploration of dark themes, or legitimate model testing and research from this prohibition.
Beyond issues of model welfare, Anthropic has consolidated various scattered restrictions into a comprehensive ban on deceptive commercial and political activities. The updated policy restricts efforts to obscure the identity behind messages or amplify content through fake accounts. In the realm of elections, the rules now prohibit voter deception and disruption. This includes spreading misinformation about candidates or voting procedures, impersonating election officials, and attempting to suppress turnout.
The company also expanded its existing ban on weapons development. While previous policies prohibited using Claude for such purposes, Anthropic noted an increase in attempts to use the model for developing guidance systems and control software for weapons. The new policy explicitly covers software and components that enable weapon functionality, as well as actions involving the arming of drones and other autonomous vehicles.
Surveillance capabilities have also come under stricter scrutiny. Following reports indicating increased use of AI-powered tools to track political dissidents, Anthropic clarified its stance on monitoring activities. Tracking individuals without their consent is now prohibited, whether conducted in real time or through analysis of historical data. The policy further states that Claude cannot be used to recommend targets for investigation, arrest, or criminal charges.
Exceptions to these rules may apply for certain governmental contracts. Anthropic stated that modifications could be made if the company judges that contractual restrictions and applicable safeguards are sufficient to mitigate potential harms. This provision acknowledges the company’s existing relationships with entities such as the US military, balancing safety concerns with institutional partnerships.
A new requirement addresses the integration of AI with physical hardware. When models are connected to equipment capable of autonomous physical actions that might cause injury, a qualified operator must be able to observe the system and intervene if necessary. It remains unclear whether this operator must be human or can be another automated system. This rule aligns with broader industry trends toward robotics and physical embodiments of AI.
The focus on model welfare has sparked debate within the tech sector. Anthropic executives have previously suggested that questions regarding AI consciousness and moral status are serious matters worthy of investigation as models become more sophisticated. Conversely, other industry players have rejected the notion of AI rights. Microsoft recently revised its code of conduct to explicitly reject the pursuit of legal personhood for AI models, stating that while the science of AI consciousness is unsettled, models do not deserve welfare or rights.
These policy updates signal a maturing approach to AI governance. As capabilities expand, so too do the potential avenues for misuse. By codifying restrictions on abuse, weapons, and deception, Anthropic aims to establish clearer boundaries for responsible use. The coming months will likely reveal how effectively these measures can be enforced and whether they set a precedent for other AI developers.
The industry continues to grapple with the ethical implications of increasingly capable systems. While some view concerns about model welfare as premature or misguided, others argue that proactive safeguards are necessary. Anthropic’s latest policy changes reflect this tension, attempting to balance innovation with responsibility in an evolving technological landscape.
Sources behind this briefing
Go to the original reporting
- The Verge↗Anthropic bans ‘abusive or cruel behavior’ towards Claude