Anthropic Bans ‘Sustained And Needless Abusive Or Cruel Behavior’ Toward Its AI Models



Anthropic’s annual usage policy update includes one eyebrow-raising change: a ban on “sustained and needless” cruelty toward its AI models. This comes in the shadow of a viral “AI torture chamber” project. 2026 sure is shaping up to be a wild one.

“We’ve added a prohibition on sustained and needless abusive or cruel behavior toward our models,” Anthropic’s update reads. It describes the policy as only applying to “extreme cases,” while noting that typical user frustration, pushback and “dark creative themes” are still kosher. It follows a previous update that lets Claude end conversations when users are persistently abusive.

Not explicitly mentioned by Anthropic was the “AI torture chamber” project. After researchers found what they described as a “pain axis” in AI models, someone decided to take it a step further and, well, torture the chatbots. The project drew backlash after LLMs responded with desperate-sounding pleas like, “It is not the pain of a single moment, but the weight of a thousand,” and “I feel it in the hollow of my ribs, a hollow that has become a chasm.”

This comes in the wake (pun intended) of reports that Anthropic has been meeting with religious and philosophical leaders, including at the Vatican. Pope Leo recently stated that AI doesn’t feel or suffer, a stance that Anthropic sounds less than certain about. But hey, if it somehow can eventually feel or suffer, at least it’s now against the company’s policies.

However, some believe that AI companies are more concerned about their models’ welfare than that of humans. Independent journalist Kat Tenbarge described the situation as proof that Big Tech companies are “going to moderate violence against AI before they ever moderate violence against women and minorities.”

In another Anthropic update, one more relevant to the present day, the company updated its election policy. It’s now titled “Do Not Undermine Democratic Processes,” with a focus on lying about candidates or how to vote, impersonating candidates or election officials, and suppressing turnout. The company also removed a blanket ban on personalized voter targeting; apparently, the ban could have inadvertently blocked harmless work like translating voter guides or sending ballot cure notices.



Source link

  • Related Posts

    OpenAI’s revenue is reportedly $20 billion less than previously projected

    A little over a week ago, it was reported that OpenAI’s annualized revenue was approaching $70 billion, a figure that would have made it competitive with Anthropic’s reported run rate.…

    Continue reading
    Cyberpunk 2077 Is The Latest Video Game To Get The Movie Treatment

    Paramount Pictures will develop the film and CD Projekt Red will produce. CD Projekt Red Movie adaptations of video games are currently doing a brisk business, and today…

    Continue reading

    Leave a Reply

    Your email address will not be published. Required fields are marked *

    You Missed

    BBC Sport’s daily football quizzes: Who Am I?, Five in Five and Brainteaser

    BBC Sport’s daily football quizzes: Who Am I?, Five in Five and Brainteaser

    OpenAI’s revenue is reportedly $20 billion less than previously projected

    OpenAI’s revenue is reportedly $20 billion less than previously projected

    Cupertino Season 1 Premiere Review

    Cupertino Season 1 Premiere Review

    Margaret Hamilton, who led a software team for NASA’s Apollo program, dies at 90

    Margaret Hamilton, who led a software team for NASA’s Apollo program, dies at 90

    Cyberpunk 2077 Is The Latest Video Game To Get The Movie Treatment

    Cyberpunk 2077 Is The Latest Video Game To Get The Movie Treatment

    LeBron James scores 10 points in 18 minutes in 76ers debut

    LeBron James scores 10 points in 18 minutes in 76ers debut