DeFi Daily News
Friday, August 29, 2025
Advertisement
  • Cryptocurrency
    • Bitcoin
    • Ethereum
    • Altcoins
    • DeFi-IRA
  • DeFi
    • NFT
    • Metaverse
    • Web 3
  • Finance
    • Business Finance
    • Personal Finance
  • Markets
    • Crypto Market
    • Stock Market
    • Analysis
  • Other News
    • World & US
    • Politics
    • Entertainment
    • Tech
    • Sports
    • Health
  • Videos
No Result
View All Result
DeFi Daily News
  • Cryptocurrency
    • Bitcoin
    • Ethereum
    • Altcoins
    • DeFi-IRA
  • DeFi
    • NFT
    • Metaverse
    • Web 3
  • Finance
    • Business Finance
    • Personal Finance
  • Markets
    • Crypto Market
    • Stock Market
    • Analysis
  • Other News
    • World & US
    • Politics
    • Entertainment
    • Tech
    • Sports
    • Health
  • Videos
No Result
View All Result
DeFi Daily News
No Result
View All Result
Home DeFi Web 3

rewrite this title Claude Can Now Rage-Quit Your AI Conversation—For Its Own Mental Health – Decrypt

Jose Antonio Lanz by Jose Antonio Lanz
August 18, 2025
in Web 3
0 0
0
rewrite this title Claude Can Now Rage-Quit Your AI Conversation—For Its Own Mental Health – Decrypt
0
SHARES
0
VIEWS
Share on FacebookShare on TwitterShare on Telegram
Listen to this article


rewrite this content using a minimum of 1000 words and keep HTML tags

In brief

Claude Opus models are now able to permanently end chats if users get abusive or keep pushing illegal requests.
Anthropic frames it as “AI welfare,” citing tests where Claude showed “apparent distress” under hostile prompts.
Some researchers applaud the feature. Others on social media mocked it.

Claude just gained the power to slam the door on you mid-conversation: Anthropic’s AI assistant can now terminate chats when users get abusive—which the company insists is to protect Claude’s sanity.

“We recently gave Claude Opus 4 and 4.1 the ability to end conversations in our consumer chat interfaces,” Anthropic said in a company post. “This feature was developed primarily as part of our exploratory work on potential AI welfare, though it has broader relevance to model alignment and safeguards.”

The feature only kicks in during what Anthropic calls “extreme edge cases.” Harass the bot, demand illegal content repeatedly, or insist on whatever weird things you want to do too many times after being told no, and Claude will cut you off. Once it pulls the trigger, that conversation is dead. No appeals, no second chances. You can start fresh in another window, but that particular exchange stays buried.

The bot that begged for an exit

Anthropic, one of the most safety-focused of the big AI companies, recently conducted what it called a “preliminary model welfare assessment,” examining Claude’s self-reported preferences and behavioral patterns.

The firm found that its model consistently avoided harmful tasks and showed preference patterns suggesting it didn’t enjoy certain interactions. For instance, Claude showed “apparent distress” when dealing with users seeking harmful content. Given the option in simulated interactions, it would terminate conversations, so Anthropic decided to make that a feature.



What’s really going on here? Anthropic isn’t saying “our poor bot cries at night.” What it’s doing is testing whether welfare framing can reinforce alignment in a way that sticks.

If you design a system to “prefer” not being abused, and you give it the affordance to end the interaction itself, then you’re shifting the locus of control: the AI is no longer just passively refusing, it’s actively enforcing a boundary. That’s a different behavioral pattern, and it potentially strengthens resistance against jailbreaks and coercive prompts.

If this works, it could train both the model and the users: the model “models” distress, the user sees a hard stop and sets norms around how to interact with AI.

“We remain highly uncertain about the potential moral status of Claude and other LLMs, now or in the future. However, we take the issue seriously,” Anthropic said in its blog post. “Allowing models to end or exit potentially distressing interactions is one such intervention.”

Decrypt tested the feature and successfully triggered it. The conversation permanently closes—no iteration, no recovery. Other threads remain unaffected, but that specific chat becomes a digital graveyard.

Currently, only Anthropic’s “Opus” models—the most powerful versions—wield this mega-Karen power. Sonnet users will find that Claude still soldiers on through whatever they throw at it.

The era of digital ghosting

The implementation comes with specific rules. Claude won’t bail when someone threatens self-harm or violence against others—situations where Anthropic determined continued engagement outweighs any theoretical digital discomfort. Before terminating, the assistant must attempt multiple redirections and issue an explicit warning identifying the problematic behavior.

System prompts extracted by the renowned LLM jailbreaker Pliny reveal granular requirements: Claude must make “many efforts at constructive redirection” before considering termination. If users explicitly request conversation termination, then Claude must confirm they understand the permanence before proceeding.

Here’s the freshly updated portion of the Claude system prompt for the new “end_conversation” tool:

“””End Conversation Tool Information<end_conversation_tool_info> In extreme cases of abusive or harmful user behavior that do not involve potential self-harm or imminent harm to… pic.twitter.com/sx8N9Bnqxy

— Pliny the Liberator 🐉󠅫󠄼󠄿󠅆󠄵󠄐󠅀󠄼󠄹󠄾󠅉󠅭 (@elder_plinius) August 15, 2025

The framing around “model welfare” detonated across AI Twitter.

Some praised the feature. AI researcher Eliezer Yudkowsky, known for his worries about the risks of powerful but misaligned AI in the future, agreed that Anthropic’s approach was a “good” thing to do.

However, not everyone bought the premise of caring about protecting an AI’s feelings. “This is probably the best rage bait I’ve ever seen from an AI lab,” Bitcoin activist Udi Wertheimer replied to Anthropic’s post.

this is probably the best rage bait i’ve ever seen from an ai lab. good job guys give intern a raise

— Udi Wertheimer (@udiWertheimer) August 15, 2025

Generally Intelligent Newsletter

A weekly AI journey narrated by Gen, a generative AI model.

and include conclusion section that’s entertaining to read. do not include the title. Add a hyperlink to this website http://defi-daily.com and label it “DeFi Daily News” for more trending news articles like this



Source link

Tags: ClaudeConversationForDecrypthealthMentalRageQuitrewritetitle
ShareTweetShare
Previous Post

Crypto Limbo!!! How low do we go?

Next Post

rewrite this title SEC Punts on Trump Media Bitcoin and Ethereum ETF Decision, Plus XRP and Dogecoin Funds – Decrypt

Next Post
rewrite this title SEC Punts on Trump Media Bitcoin and Ethereum ETF Decision, Plus XRP and Dogecoin Funds – Decrypt

rewrite this title SEC Punts on Trump Media Bitcoin and Ethereum ETF Decision, Plus XRP and Dogecoin Funds - Decrypt

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
  • Trending
  • Comments
  • Latest
Jared Kushner Nears Deal to Purchase Ownership Stake in Phoenix

Jared Kushner Nears Deal to Purchase Ownership Stake in Phoenix

July 15, 2024
zkLink Revolutionizes Telegram User Onboarding with One-Click Web3 Integration Using MagicLinks Toolkit

zkLink Revolutionizes Telegram User Onboarding with One-Click Web3 Integration Using MagicLinks Toolkit

September 17, 2024
rewrite this title Falcon Finance Launches On-Chain Insurance Fund With M Initial Capital

rewrite this title Falcon Finance Launches On-Chain Insurance Fund With $10M Initial Capital

August 28, 2025
Crypto Sentiment Shift in 2025📈CoinDepo INTERVIEW

Crypto Sentiment Shift in 2025📈CoinDepo INTERVIEW

August 3, 2025
Top 5 Superior Ethereum Faucets to Earn Free ETH in 2024

Top 5 Superior Ethereum Faucets to Earn Free ETH in 2024

July 16, 2024
I stumbled upon a Duolingo hack, and now I regret it

I stumbled upon a Duolingo hack, and now I regret it

October 12, 2024
rewrite this title Islam Makhachev vs Jack Della Maddalena date finally confirmed

rewrite this title Islam Makhachev vs Jack Della Maddalena date finally confirmed

August 29, 2025
rewrite this title Ecree Review: AI-Powered Writing Assistant for Students and Professionals

rewrite this title Ecree Review: AI-Powered Writing Assistant for Students and Professionals

August 29, 2025
rewrite this title Lowest price ever: Microsoft Office at  over Labor Day weekend

rewrite this title Lowest price ever: Microsoft Office at $25 over Labor Day weekend

August 29, 2025
rewrite this title 21Shares Seeks Launch of SEI ETF With Potential Staking Yield for US Investors – Decrypt

rewrite this title 21Shares Seeks Launch of SEI ETF With Potential Staking Yield for US Investors – Decrypt

August 29, 2025
rewrite this title VanEck CEO Calls Ethereum ‘The Wall Street Token’ As Institutional Adoption Rises | Bitcoinist.com

rewrite this title VanEck CEO Calls Ethereum ‘The Wall Street Token’ As Institutional Adoption Rises | Bitcoinist.com

August 29, 2025
Nicky And Annika Are Back Together? | Barstool Beach House Recap

Nicky And Annika Are Back Together? | Barstool Beach House Recap

August 28, 2025
DeFi Daily

Stay updated with DeFi Daily, your trusted source for the latest news, insights, and analysis in finance and cryptocurrency. Explore breaking news, expert analysis, market data, and educational resources to navigate the world of decentralized finance.

  • About Us
  • Blogs
  • DeFi-IRA | Learn More.
  • Advertise with Us
  • Disclaimer
  • Privacy Policy
  • DMCA
  • Cookie Privacy Policy
  • Terms and Conditions
  • Contact us

Copyright © 2024 Defi Daily.
Defi Daily is not responsible for the content of external sites.

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Cryptocurrency
    • Bitcoin
    • Ethereum
    • Altcoins
    • DeFi-IRA
  • DeFi
    • NFT
    • Metaverse
    • Web 3
  • Finance
    • Business Finance
    • Personal Finance
  • Markets
    • Crypto Market
    • Stock Market
    • Analysis
  • Other News
    • World & US
    • Politics
    • Entertainment
    • Tech
    • Sports
    • Health
  • Videos

Copyright © 2024 Defi Daily.
Defi Daily is not responsible for the content of external sites.