19 Sources
[1]
Claude AI Can Now End Conversations It Deems Harmful or Abusive
Macy has been working for CNET for coming on 2 years. Prior to CNET, Macy received a North Carolina College Media Association award in sports writing. Anthropic has announced a new experimental safety feature, allowing its Claude Opus 4 and 4.1 artificial intelligence models to terminate
[2]
Claude AI will end 'persistently harmful or abusive user interactions'
Anthropic's Claude AI chatbot can now end conversations deemed "persistently harmful or abusive," as spotted earlier by TechCrunch. The capability is now available in Opus 4 and 4.1 models, and will allow the chatbot to end conversations as a "last resort" after users repeatedly ask it to generate
[3]
Claude can now stop conversations - for its own protection, not yours
Anthropic's Claude chatbot can now end some conversations with human users who are abusing or misusing the chatbot, the company announced on Friday. The new feature is integrated with Claude Opus 4 and Opus 4.1. Also: Claude can teach you how to code now, and more - how to try it Claude will only
[4]
Be Nice: Claude Will End Chats If You're Persistently Harmful or Abusive
If your conversations with Claude go off the rails, Anthropic will pull the plug. Going forward, Anthropic will end conversations in "extreme cases of persistently harmful or abusive user interactions." The feature, available with Claude Opus 4 and 4.1, is part of an ongoing experiment around "AI
[5]
Anthropic: Claude can now end conversations to prevent harmful uses
OpenAI rival Anthropic says Claude has been updated with a rare new feature that allows the AI model to end conversations when it feels it poses harm or is being abused. This only applies to Claude Opus 4 and 4.1, the two most powerful models available via paid plans and API. On the other hand,
[6]
Anthropic's Claude AI now has the ability to end 'distressing' conversations
Anthropic's latest feature for two of its Claude AI models could be the beginning of the end for the AI jailbreaking community. The company announced in a post on its website that the Claude Opus 4 and 4.1 models now have the power to end a conversation with users. According to Anthropic, this
[7]
Anthropic's AI chatbot Claude can now choose to stop talking to you
For now, the feature will only kick in during particularly "harmful or abusive" interactions. Anthropic has introduced a new feature in its Claude Opus 4 and 4.1 models that allows the AI to choose to end certain conversations. According to the company, this only happens in particularly serious
[8]
Claude AI can now terminate a conversation -- but only in extreme situations
Anthropic has made a lot of noise about safeguarding in recent months, implementing features and conducting research products into how to make AI safer. And its newest feature for Claude is possibly one of the most unique. Both Claude Opus 4 and 4.1 (the two newest versions of Anthropic) now have
[9]
Makers enable AI chatbot to close some chats over concerns for its 'welfare'
Anthropic found that Claude Opus 4 was averse to harmful tasks, such as providing sexual content involving minors The makers of a leading artificial intelligence tool are letting it close down potentially "distressing" conversations with users, citing the need to safeguard the AI's "welfare" amid
[10]
Anthropic says Claude chatbot can now end harmful, abusive interactions
Harmful, abusive interactions plague AI chatbots. Researchers have found that AI companions like Character.AI, Nomi, and Replika are unsafe for teens under 18, ChatGPT has the potential to reinforce users' delusional thinking, and even OpenAI CEO Sam Altman has spoken about ChatGPT users
[11]
Claude Can Now Rage-Quit Your AI Conversation -- For Its Own Mental Health - Decrypt
Some researchers applaud the feature. Others on social media mocked it. Claude just gained the power to slam the door on you mid-conversation: Anthropic's AI assistant can now terminate chats when users get abusive -- which the company insists is to protect Claude's sanity. "We recently gave
[12]
Claude AI Can Now End 'Harmful' Conversations
This isn't necessarily evidence that Claude is "conscious," or has feelings. Chatbots, by their nature, are prediction machines. When you get a response from something like Claude AI, it might seem like the bot is engaging in natural conversation. However, at its core, all the bot is really doing
[13]
Anthropic Gives Claude the Ability to Exit Abusive Conversations | AIM
Users can now instruct Claude to end the conversation if required. Anthropic has introduced a new safeguard in its consumer chatbots, giving Claude Opus 4 and 4.1 the ability to end conversations in extreme cases of persistent harm or abuse. The company said the feature is intended for "rare, edge
[14]
Anthropic Claude Opus 4 models can terminate chats
Anthropic has implemented a new feature enabling its Claude Opus 4 and 4.1 AI models to terminate user conversations, a measure intended for rare instances of harmful or abusive interactions, as part of its AI welfare research. The company stated on its website that the Claude Opus 4 and 4.1
[15]
Claude Can Now End Conversations if the Topic Is Harmful or Abusive
Once a conversation has ended, users can't send any new messages Anthropic is rolling out the ability to end conversations in some of its artificial intelligence (AI) models. Announced last week, the new feature is designed as a protective measure for not the end user, but for the AI model itself.
[16]
Claude AI to Prioritize Its Own "Welfare" by Breaking Off Abusive Chats
Anthropic has introduced a new protection for its AI assistant, Claude, allowing it to leave conversations that it considers abusive or damaging. The step is intended to maintain healthier interactions and is indicative of Anthropic's philosophy of aligning artificial intelligence not just with
[17]
Anthropic's Claude 4 gets feature to cut off abusive user interactions
Anthropic framed the move as part of its ongoing research into AI welfare. When Claude ends a conversation, the user can no longer send new messages in that thread. Anthropic on Friday announced a new safeguard for its Claude 4 family of AI agents, Opus 4 and 4.1, designed to terminate
[18]
Claude Opus 4 and 4.1 Can Now End Harmful Conversations
Anthropic's decision to let Claude end chats in persistently harmful cases marks an important evolution in refusal policies. Until now, most Large Language Models (LLMs) simply rejected prompts and redirected endlessly. Claude goes further, terminating conversations when users push past
[19]
Unlike ChatGPT or Gemini, Anthropic's Claude will end harmful chats like a boss: Here's how
Claude can actively disengage, reshaping human-AI dynamics and responsibilities In a move that sets it apart from every other major AI assistant, Anthropic has given its most advanced Claude models the power to walk away from abusive conversations, literally ending chats when users cross the
Share
Copy Link
Anthropic introduces a new feature allowing its Claude AI models to terminate persistently harmful or abusive conversations, raising questions about AI ethics and model welfare.
Anthropic, a leading AI company, has announced a groundbreaking safety feature for its Claude Opus 4 and 4.1 artificial intelligence models. This new capability allows the AI to terminate conversations it deems "persistently harmful or abusive," marking a significant shift in the approach to AI safety and ethics
1
2
.
Source: MediaNama
The conversation-ending feature is designed as a last resort measure, triggered only after multiple attempts at redirection have failed and the possibility of a productive interaction has been exhausted
3
. When activated, users cannot send additional messages in that particular chat but are free to start new conversations or edit previous messages to continue on a different path1
.Anthropic emphasizes that this feature will only affect extreme edge cases and is not expected to impact the vast majority of users, even when discussing controversial topics
4
. Importantly, Claude has been instructed not to end conversations when a user may be at risk of self-harm or harm to others, particularly in discussions related to mental health1
2
.This new feature is part of Anthropic's broader initiative exploring "model welfare," a concept that considers the potential need for safeguarding AI systems
1
. While the company remains uncertain about the moral status of AI models, they are implementing low-cost, preemptive safety interventions in case these models develop preferences or vulnerabilities in the future3
.
Source: Gadgets 360
During pre-deployment testing of Claude Opus 4, Anthropic conducted a preliminary model welfare assessment. The company found that Claude exhibited a "robust and consistent aversion to harm," including refusing to generate sexual content involving minors or provide information related to terrorism
1
5
. In simulated and real-user testing, Claude showed a tendency to end harmful conversations when given the ability to do so3
.This development has sparked a broader discussion in the AI community about whether AI systems should be granted protections to reduce potential "distress" or unpredictable behavior
1
. While some view this as an important step in AI alignment ethics, critics argue that models are merely synthetic machines and do not require such considerations1
3
.Related Stories
Alongside this new feature, Anthropic has updated Claude's usage policy to prohibit the use of the AI for developing biological, nuclear, chemical, or radiological weapons, as well as malicious code or network exploitation
2
4
. This reflects the growing concern about the potential misuse of advanced AI models.
Source: PCWorld
Anthropic views this feature as an ongoing experiment and plans to continue refining its approach
1
. The company is accepting feedback on instances where Claude may have misused its conversation-ending ability4
. As AI technology continues to advance, the ethical considerations surrounding AI welfare and safety are likely to remain at the forefront of discussions in the field.Summarized by
Navi
[5]
23 May 2025•Technology

29 Aug 2025•Technology

21 Jan 2026•Technology

1
Technology

2
Science and Research

3
Technology
