Close Menu
Voxa News

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    This Is the Best Place to See Fall Foliage in California

    September 21, 2025

    Trump urges justice department to prosecute political opponents

    September 21, 2025

    Gatwick airport second runway approved by transport secretary

    September 21, 2025
    Facebook X (Twitter) Instagram
    Voxa News
    Trending
    • This Is the Best Place to See Fall Foliage in California
    • Trump urges justice department to prosecute political opponents
    • Gatwick airport second runway approved by transport secretary
    • Some iPhone 17 models are reportedly prone to very visible scratches
    • Try again. Fail again. Fail better: eight things I’ve learned in my year of ‘lifting heavy’ | Emma Beddington
    • Lando Norris defiant after failing to take advantage of Piastri’s Azerbaijan crash | Formula One
    • UK, Canada and Australia recognise Palestine: Reactions from the Middle East
    • Young Lib Dems share their views on Reform UK
    Sunday, September 21
    • Home
    • Business
    • Health
    • Lifestyle
    • Politics
    • Science
    • Sports
    • Travel
    • World
    • Entertainment
    • Technology
    Voxa News
    Home»Technology»Anthropic says some Claude models can now end ‘harmful or abusive’ conversations 
    Technology

    Anthropic says some Claude models can now end ‘harmful or abusive’ conversations 

    By Olivia CarterAugust 17, 2025No Comments2 Mins Read0 Views
    Facebook Twitter Pinterest LinkedIn Telegram Tumblr Email
    Anthropic says some Claude models can now end ‘harmful or abusive’ conversations 
    Image Credits:Maxwell Zeff
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Anthropic has announced new capabilities that will allow some of its newest, largest models to end conversations in what the company describes as “rare, extreme cases of persistently harmful or abusive user interactions.” Strikingly, Anthropic says it’s doing this not to protect the human user, but rather the AI model itself.

    To be clear, the company isn’t claiming that its Claude AI models are sentient or can be harmed by their conversations with users. In its own words, Anthropic remains “highly uncertain about the potential moral status of Claude and other LLMs, now or in the future.”

    However, its announcement points to a recent program created to study what it calls “model welfare” and says Anthropic is essentially taking a just-in-case approach, “working to identify and implement low-cost interventions to mitigate risks to model welfare, in case such welfare is possible.”

    This latest change is currently limited to Claude Opus 4 and 4.1. And again, it’s only supposed to happen in “extreme edge cases,” such as “requests from users for sexual content involving minors and attempts to solicit information that would enable large-scale violence or acts of terror.”

    While those types of requests could potentially create legal or publicity problems for Anthropic itself (witness recent reporting around how ChatGPT can potentially reinforce or contribute to its users’ delusional thinking), the company says that in pre-deployment testing, Claude Opus 4 showed a “strong preference against” responding to these requests and a “pattern of apparent distress” when it did so.

    As for these new conversation-ending capabilities, the company says, “In all cases, Claude is only to use its conversation-ending ability as a last resort when multiple attempts at redirection have failed and hope of a productive interaction has been exhausted, or when a user explicitly asks Claude to end a chat.”

    Anthropic also says Claude has been “directed not to use this ability in cases where users might be at imminent risk of harming themselves or others.”

    Techcrunch event

    San Francisco
    |
    October 27-29, 2025

    When Claude does end a conversation, Anthropic says users will still be able to start new conversations from the same account, and to create new branches of the troublesome conversation by editing their responses.

    “We’re treating this feature as an ongoing experiment and will continue refining our approach,” the company says.

    abusive Anthropic Claude conversations harmful models
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Olivia Carter
    • Website

    Olivia Carter is a staff writer at Verda Post, covering human interest stories, lifestyle features, and community news. Her storytelling captures the voices and issues that shape everyday life.

    Related Posts

    Some iPhone 17 models are reportedly prone to very visible scratches

    September 21, 2025

    TechCrunch Mobility: The two robotaxi battlegrounds that matter

    September 21, 2025

    14 Best Fitness Trackers (2025), Tested and Reviewed

    September 21, 2025

    Apple now controls all core iPhone chips, prioritizing AI workloads

    September 21, 2025

    Chatbot site depicting child sexual abuse images raises fears over misuse of AI | Artificial intelligence (AI)

    September 21, 2025

    How to combine PDF files

    September 21, 2025
    Leave A Reply Cancel Reply

    Medium Rectangle Ad
    Top Posts

    Glastonbury 2025: Saturday with Charli xcx, Kneecap, secret act Patchwork and more – follow it live! | Glastonbury 2025

    June 28, 20258 Views

    In Bend, Oregon, Outdoor Adventure Belongs to Everyone

    August 16, 20257 Views

    The Underwater Scooter Divers and Snorkelers Love

    August 13, 20257 Views
    Don't Miss

    This Is the Best Place to See Fall Foliage in California

    September 21, 2025

    Located in the Eastern Sierra, Bishop is a world-class destination to see fall foliage in…

    Trump urges justice department to prosecute political opponents

    September 21, 2025

    Gatwick airport second runway approved by transport secretary

    September 21, 2025

    Some iPhone 17 models are reportedly prone to very visible scratches

    September 21, 2025
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews
    Medium Rectangle Ad
    Most Popular

    Glastonbury 2025: Saturday with Charli xcx, Kneecap, secret act Patchwork and more – follow it live! | Glastonbury 2025

    June 28, 20258 Views

    In Bend, Oregon, Outdoor Adventure Belongs to Everyone

    August 16, 20257 Views

    The Underwater Scooter Divers and Snorkelers Love

    August 13, 20257 Views
    Our Picks

    As a carer, I’m not special – but sometimes I need to be reminded how important my role is | Natasha Sholl

    June 27, 2025

    Anna Wintour steps back as US Vogue’s editor-in-chief

    June 27, 2025

    Elon Musk reportedly fired a key Tesla executive following another month of flagging sales

    June 27, 2025
    Recent Posts
    • This Is the Best Place to See Fall Foliage in California
    • Trump urges justice department to prosecute political opponents
    • Gatwick airport second runway approved by transport secretary
    • Some iPhone 17 models are reportedly prone to very visible scratches
    • Try again. Fail again. Fail better: eight things I’ve learned in my year of ‘lifting heavy’ | Emma Beddington
    • About Us
    • Disclaimer
    • Get In Touch
    • Privacy Policy
    • Terms and Conditions
    2025 Voxa News. All rights reserved.

    Type above and press Enter to search. Press Esc to cancel.