The post Claude Can Now Rage-Quit Your AI Conversation—For Its Own Mental Health appeared on BitcoinEthereumNews.com. In brief Claude Opus models are now able to permanently end chats if users get abusive or keep pushing illegal requests. Anthropic frames it as “AI welfare,” citing tests where Claude showed “apparent distress” under hostile prompts. Some researchers applaud the feature. Others on social media mocked it. Claude just gained the power to slam the door on you mid-conversation: Anthropic’s AI assistant can now terminate chats when users get abusive—which the company insists is to protect Claude’s sanity. “We recently gave Claude Opus 4 and 4.1 the ability to end conversations in our consumer chat interfaces,” Anthropic said in a company post. “This feature was developed primarily as part of our exploratory work on potential AI welfare, though it has broader relevance to model alignment and safeguards.” The feature only kicks in during what Anthropic calls “extreme edge cases.” Harass the bot, demand illegal content repeatedly, or insist on whatever weird things you want to do too many times after being told no, and Claude will cut you off. Once it pulls the trigger, that conversation is dead. No appeals, no second chances. You can start fresh in another window, but that particular exchange stays buried. The bot that begged for an exit Anthropic, one of the most safety-focused of the big AI companies, recently conducted what it called a “preliminary model welfare assessment,” examining Claude’s self-reported preferences and behavioral patterns. The firm found that its model consistently avoided harmful tasks and showed preference patterns suggesting it didn’t enjoy certain interactions. For instance, Claude showed “apparent distress” when dealing with users seeking harmful content. Given the option in simulated interactions, it would terminate conversations, so Anthropic decided to make that a feature.  What’s really going on here? Anthropic isn’t saying “our poor bot cries at night.” What it’s… The post Claude Can Now Rage-Quit Your AI Conversation—For Its Own Mental Health appeared on BitcoinEthereumNews.com. In brief Claude Opus models are now able to permanently end chats if users get abusive or keep pushing illegal requests. Anthropic frames it as “AI welfare,” citing tests where Claude showed “apparent distress” under hostile prompts. Some researchers applaud the feature. Others on social media mocked it. Claude just gained the power to slam the door on you mid-conversation: Anthropic’s AI assistant can now terminate chats when users get abusive—which the company insists is to protect Claude’s sanity. “We recently gave Claude Opus 4 and 4.1 the ability to end conversations in our consumer chat interfaces,” Anthropic said in a company post. “This feature was developed primarily as part of our exploratory work on potential AI welfare, though it has broader relevance to model alignment and safeguards.” The feature only kicks in during what Anthropic calls “extreme edge cases.” Harass the bot, demand illegal content repeatedly, or insist on whatever weird things you want to do too many times after being told no, and Claude will cut you off. Once it pulls the trigger, that conversation is dead. No appeals, no second chances. You can start fresh in another window, but that particular exchange stays buried. The bot that begged for an exit Anthropic, one of the most safety-focused of the big AI companies, recently conducted what it called a “preliminary model welfare assessment,” examining Claude’s self-reported preferences and behavioral patterns. The firm found that its model consistently avoided harmful tasks and showed preference patterns suggesting it didn’t enjoy certain interactions. For instance, Claude showed “apparent distress” when dealing with users seeking harmful content. Given the option in simulated interactions, it would terminate conversations, so Anthropic decided to make that a feature.  What’s really going on here? Anthropic isn’t saying “our poor bot cries at night.” What it’s…

Claude Can Now Rage-Quit Your AI Conversation—For Its Own Mental Health

2025/08/19 11:43
Okuma süresi: 4 dk
Bu içerikle ilgili geri bildirim veya endişeleriniz için lütfen [email protected] üzerinden bizimle iletişime geçin.

In brief

  • Claude Opus models are now able to permanently end chats if users get abusive or keep pushing illegal requests.
  • Anthropic frames it as “AI welfare,” citing tests where Claude showed “apparent distress” under hostile prompts.
  • Some researchers applaud the feature. Others on social media mocked it.

Claude just gained the power to slam the door on you mid-conversation: Anthropic’s AI assistant can now terminate chats when users get abusive—which the company insists is to protect Claude’s sanity.

“We recently gave Claude Opus 4 and 4.1 the ability to end conversations in our consumer chat interfaces,” Anthropic said in a company post. “This feature was developed primarily as part of our exploratory work on potential AI welfare, though it has broader relevance to model alignment and safeguards.”

The feature only kicks in during what Anthropic calls “extreme edge cases.” Harass the bot, demand illegal content repeatedly, or insist on whatever weird things you want to do too many times after being told no, and Claude will cut you off. Once it pulls the trigger, that conversation is dead. No appeals, no second chances. You can start fresh in another window, but that particular exchange stays buried.

The bot that begged for an exit

Anthropic, one of the most safety-focused of the big AI companies, recently conducted what it called a “preliminary model welfare assessment,” examining Claude’s self-reported preferences and behavioral patterns.

The firm found that its model consistently avoided harmful tasks and showed preference patterns suggesting it didn’t enjoy certain interactions. For instance, Claude showed “apparent distress” when dealing with users seeking harmful content. Given the option in simulated interactions, it would terminate conversations, so Anthropic decided to make that a feature.

What’s really going on here? Anthropic isn’t saying “our poor bot cries at night.” What it’s doing is testing whether welfare framing can reinforce alignment in a way that sticks.

If you design a system to “prefer” not being abused, and you give it the affordance to end the interaction itself, then you’re shifting the locus of control: the AI is no longer just passively refusing, it’s actively enforcing a boundary. That’s a different behavioral pattern, and it potentially strengthens resistance against jailbreaks and coercive prompts.

If this works, it could train both the model and the users: the model “models” distress, the user sees a hard stop and sets norms around how to interact with AI.

“We remain highly uncertain about the potential moral status of Claude and other LLMs, now or in the future. However, we take the issue seriously,” Anthropic said in its blog post. “Allowing models to end or exit potentially distressing interactions is one such intervention.”

Decrypt tested the feature and successfully triggered it. The conversation permanently closes—no iteration, no recovery. Other threads remain unaffected, but that specific chat becomes a digital graveyard.

Currently, only Anthropic’s “Opus” models—the most powerful versions—wield this mega-Karen power. Sonnet users will find that Claude still soldiers on through whatever they throw at it.

The era of digital ghosting

The implementation comes with specific rules. Claude won’t bail when someone threatens self-harm or violence against others—situations where Anthropic determined continued engagement outweighs any theoretical digital discomfort. Before terminating, the assistant must attempt multiple redirections and issue an explicit warning identifying the problematic behavior.

System prompts extracted by the renowned LLM jailbreaker Pliny reveal granular requirements: Claude must make “many efforts at constructive redirection” before considering termination. If users explicitly request conversation termination, then Claude must confirm they understand the permanence before proceeding.

The framing around “model welfare” detonated across AI Twitter.

Some praised the feature. AI researcher Eliezer Yudkowsky, known for his worries about the risks of powerful but misaligned AI in the future, agreed that Anthropic’s approach was a “good” thing to do.

However, not everyone bought the premise of caring about protecting an AI’s feelings. “This is probably the best rage bait I’ve ever seen from an AI lab,” Bitcoin activist Udi Wertheimer replied to Anthropic’s post.

Generally Intelligent Newsletter

A weekly AI journey narrated by Gen, a generative AI model.

Source: https://decrypt.co/335732/claude-rage-quit-conversation-own-mental-health

Piyasa Fırsatı
Threshold Logosu
Threshold Fiyatı(T)
$0.007085
$0.007085$0.007085
+2.03%
USD
Threshold (T) Canlı Fiyat Grafiği
Sorumluluk Reddi: Bu sitede yeniden yayınlanan makaleler, halka açık platformlardan alınmıştır ve yalnızca bilgilendirme amaçlıdır. MEXC'nin görüşlerini yansıtmayabilir. Tüm hakları telif sahiplerine aittir. Herhangi bir içeriğin üçüncü taraf haklarını ihlal ettiğini düşünüyorsanız, kaldırılması için lütfen [email protected] ile iletişime geçin. MEXC, içeriğin doğruluğu, eksiksizliği veya güncelliği konusunda hiçbir garanti vermez ve sağlanan bilgilere dayalı olarak alınan herhangi bir eylemden sorumlu değildir. İçerik, finansal, yasal veya diğer profesyonel tavsiye niteliğinde değildir ve MEXC tarafından bir tavsiye veya onay olarak değerlendirilmemelidir.

Ayrıca Şunları da Beğenebilirsiniz

CME Group to launch options on XRP and SOL futures

CME Group to launch options on XRP and SOL futures

The post CME Group to launch options on XRP and SOL futures appeared on BitcoinEthereumNews.com. CME Group will offer options based on the derivative markets on Solana (SOL) and XRP. The new markets will open on October 13, after regulatory approval.  CME Group will expand its crypto products with options on the futures markets of Solana (SOL) and XRP. The futures market will start on October 13, after regulatory review and approval.  The options will allow the trading of MicroSol, XRP, and MicroXRP futures, with expiry dates available every business day, monthly, and quarterly. The new products will be added to the existing BTC and ETH options markets. ‘The launch of these options contracts builds on the significant growth and increasing liquidity we have seen across our suite of Solana and XRP futures,’ said Giovanni Vicioso, CME Group Global Head of Cryptocurrency Products. The options contracts will have two main sizes, tracking the futures contracts. The new market will be suitable for sophisticated institutional traders, as well as active individual traders. The addition of options markets singles out XRP and SOL as liquid enough to offer the potential to bet on a market direction.  The options on futures arrive a few months after the launch of SOL futures. Both SOL and XRP had peak volumes in August, though XRP activity has slowed down in September. XRP and SOL options to tap both institutions and active traders Crypto options are one of the indicators of market attitudes, with XRP and SOL receiving a new way to gauge sentiment. The contracts will be supported by the Cumberland team.  ‘As one of the biggest liquidity providers in the ecosystem, the Cumberland team is excited to support CME Group’s continued expansion of crypto offerings,’ said Roman Makarov, Head of Cumberland Options Trading at DRW. ‘The launch of options on Solana and XRP futures is the latest example of the…
Paylaş
BitcoinEthereumNews2025/09/18 00:56
Health Insurers To Cover Covid Vaccines Despite RFK, Jr. Moves

Health Insurers To Cover Covid Vaccines Despite RFK, Jr. Moves

The post Health Insurers To Cover Covid Vaccines Despite RFK, Jr. Moves appeared on BitcoinEthereumNews.com. The nation’s biggest health insurance companies will continue to cover vaccinations – including those against Covid-19 and seasonal flu – previously recommended by a federal advisory committee, America’s Health Insurance Plans said Wednesday, Sept. 17, 2025. In this photo is a free flu and Covid-19 vaccine shots available sign, CVS, Queens, New York. (Photo by: Lindsey Nicholson/Universal Images Group via Getty Images) UCG/Universal Images Group via Getty Images The nation’s biggest health insurance companies will continue to cover vaccinations – including those against Covid-19 and seasonal flu – previously recommended by a federal advisory committee. The announcement by America’s Health Insurance Plans (AHIP), which includes CVS Health’s Aetna, Humana, Cigna, Centene and an array of Blue Cross and Blue Shield plans as members, comes ahead of the first meeting of the reconstituted Advisory Committee on Immunization Practices, which now has new members chosen by U.S. Health and Human Services Secretary Robert F. Kennedy Jr., a vaccine critic. “Health plans are committed to maintaining and ensuring affordable access to vaccines,” AHIP said in a statement Wednesday. “Health plan coverage decisions for immunizations are grounded in each plan’s ongoing, rigorous review of scientific and clinical evidence, and continual evaluation of multiple sources of data.” The move by AHIP is good news for millions of Americans at a time of year when they flock to drugstores, pharmacies, physician’s offices and outpatient clinics to get their seasonal flu and Covid shots. Kennedy’s changes to U.S. vaccine policy have created confusion across the country over whether certain vaccines long covered by insurance would continue to be. AHIP has now provided some clarity for millions of Americans. “Health plans will continue to cover all ACIP-recommended immunizations that were recommended as of September 1, 2025, including updated formulations of the COVID-19 and influenza vaccines, with no cost-sharing…
Paylaş
BitcoinEthereumNews2025/09/18 03:11
US, UK, Canada Launch Operation Atlantic to Tackle Crypto Scams

US, UK, Canada Launch Operation Atlantic to Tackle Crypto Scams

Law enforcement agencies from the United States, United Kingdom, and Canada have launched Operation Atlantic, a joint effort to combat rising crypto scams and protect
Paylaş
Coinlaw2026/03/17 22:11