Close Menu
  • Home
  • AI
  • Education
  • Entertainment
  • Food Health
  • Health
  • Sports
  • Tech
  • Well Being

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

What's Hot

Inside ChatGPT’s Slow-Motion Advertising Rollout

March 11, 2026

I Bombed My First Google Interview and Got in 3 Years Later

March 11, 2026

William Shatner Says He Turned $42 From Musk Into $200K for Charity

March 11, 2026
Facebook X (Twitter) Instagram
  • Home
  • About Us
  • Advertise With Us
  • Contact us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
Facebook X (Twitter) Instagram
IQ Times Media – Smart News for a Smarter YouIQ Times Media – Smart News for a Smarter You
  • Home
  • AI
  • Education
  • Entertainment
  • Food Health
  • Health
  • Sports
  • Tech
  • Well Being
IQ Times Media – Smart News for a Smarter YouIQ Times Media – Smart News for a Smarter You
Home » Anthropic says some Claude models can now end ‘harmful or abusive’ conversations 
AI

Anthropic says some Claude models can now end ‘harmful or abusive’ conversations 

IQ TIMES MEDIABy IQ TIMES MEDIAAugust 16, 2025No Comments2 Mins Read
Facebook Twitter Pinterest LinkedIn Tumblr Email
Share
Facebook Twitter LinkedIn Pinterest Email


Anthropic has announced new capabilities that will allow some of its newest, largest models to end conversations in what the company describes as “rare, extreme cases of persistently harmful or abusive user interactions.” Strikingly, Anthropic says it’s doing this not to protect the human user, but rather the AI model itself.

To be clear, the company isn’t claiming that its Claude AI models are sentient or can be harmed by their conversations with users. In its own words, Anthropic remains “highly uncertain about the potential moral status of Claude and other LLMs, now or in the future.”

However, its announcement points to a recent program created to study what it calls “model welfare” and says Anthropic is essentially taking a just-in-case approach, “working to identify and implement low-cost interventions to mitigate risks to model welfare, in case such welfare is possible.”

This latest change is currently limited to Claude Opus 4 and 4.1. And again, it’s only supposed to happen in “extreme edge cases,” such as “requests from users for sexual content involving minors and attempts to solicit information that would enable large-scale violence or acts of terror.”

While those types of requests could potentially create legal or publicity problems for Anthropic itself (witness recent reporting around how ChatGPT can potentially reinforce or contribute to its users’ delusional thinking), the company says that in pre-deployment testing, Claude Opus 4 showed a “strong preference against” responding to these requests and a “pattern of apparent distress” when it did so.

As for these new conversation-ending capabilities, the company says, “In all cases, Claude is only to use its conversation-ending ability as a last resort when multiple attempts at redirection have failed and hope of a productive interaction has been exhausted, or when a user explicitly asks Claude to end a chat.”

Anthropic also says Claude has been “directed not to use this ability in cases where users might be at imminent risk of harming themselves or others.”

Techcrunch event

San Francisco
|
October 27-29, 2025

When Claude does end a conversation, Anthropic says users will still be able to start new conversations from the same account, and to create new branches of the troublesome conversation by editing their responses.

“We’re treating this feature as an ongoing experiment and will continue refining our approach,” the company says.



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
IQ TIMES MEDIA
  • Website

Related Posts

Google brings Gemini in Chrome to India

March 11, 2026

Amazon launches its healthcare AI assistant on its website and app

March 10, 2026

AI-powered apps struggle with long-term retention, new report shows

March 10, 2026
Add A Comment
Leave A Reply Cancel Reply

Editors Picks

Islamic boarding school in Senegal is at the center of abuse allegations

March 11, 2026

Judge to decide on scope of federal subpoena in probe of antisemitism at Penn

March 10, 2026

A Maine educator didn’t have a curriculum to teach a foundational reading skill. So she created one

March 10, 2026

Did anybody do the reading? Colleges grapple with a generational shift in learning — plus AI

March 10, 2026
Education

Islamic boarding school in Senegal is at the center of abuse allegations

By IQ TIMES MEDIAMarch 11, 20260

DAKAR, Senegal (AP) — The American Dara Academy in Senegal marketed itself to families in…

Judge to decide on scope of federal subpoena in probe of antisemitism at Penn

March 10, 2026

A Maine educator didn’t have a curriculum to teach a foundational reading skill. So she created one

March 10, 2026

Did anybody do the reading? Colleges grapple with a generational shift in learning — plus AI

March 10, 2026
IQ Times Media – Smart News for a Smarter You
Facebook X (Twitter) Instagram Pinterest Vimeo YouTube
  • Home
  • About Us
  • Advertise With Us
  • Contact us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
© 2026 iqtimes. Designed by iqtimes.

Type above and press Enter to search. Press Esc to cancel.