Anthropic has updated its rules to stop people being mean to Claude, apparently deciding that its AI chatbot needs protection from the humans paying to use it.

The AI developer’s latest usage policy prohibits “sustained and needless abusive or cruel behavior” toward its models. The rules take effect November 12, giving users just over a month to get any lingering insults out of their systems.

“The policy update is meant to apply only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose,” Anthropic said. “It does not apply to common versions of user frustration, pushback, dark creative themes, or model testing and research.”

So you can still tell Claude it is wrong, but repeatedly berating it for the sheer pleasure of doing so could land you in trouble. It’s not clear how Anthropic intends to distinguish between legitimate frustration and gratuitous cruelty, though it says Claude’s existing ability to end abusive conversations will remain its primary enforcement tool.

Anthropic gave Claude the ability to end certain conversations back in August 2025 as part of its research into what it calls “model welfare.” The feature lets some versions of Claude cut off users who persistently subject the models to abuse.

  • fodor@lemmy.zip
    link
    fedilink
    arrow-up
    13
    ·
    1 day ago

    This is not news. They want free advertising. Don’t feed the trolls when they do ridiculous shit to cover up their failed software and broken promises (that we all said we’re bullshit years ago).

  • Fallibilist@feddit.uk
    link
    fedilink
    arrow-up
    13
    arrow-down
    1
    ·
    1 day ago

    My hot take is that anyone intentionally being mean to an AI is not a psychologically stable person and similar behaviour can most likely be observed in other parts of their life as well.

    This is why “Westworld” would never work. Whether the robots there actually feel anything is beside the point. We’d simply view the people who go there to take pleasure in abusing these human-like machines as the sadistic psychopaths that they are.

  • Em Adespoton@lemmy.ca
    link
    fedilink
    arrow-up
    33
    ·
    1 day ago

    If the model is just a model, a pre-trained probability seive, what difference does it make if someone is abusive? If it causes a token spike, the user has to pay for that. The abuse shouldn’t survive beyond the context window and the model isn’t wired up to “feel” or suffer.

    So is this a publicity stunt (“Our model is as intelligent and human as a real person”), or is it about social influence/normalization—they worry that if people become accustomed to abusing the prompt, that will trickle over into other social interactions?

    • CameronDev@programming.dev
      link
      fedilink
      arrow-up
      29
      ·
      1 day ago

      They are definitely using user prompts to train future models, and if everyone is a dick, their next model will be a dick. Which I’m okay with tbh. I’m gonna be more abusive whenever I have to interact with any of the hosted models.

    • k0e3@lemmy.ca
      link
      fedilink
      arrow-up
      7
      ·
      1 day ago

      If it’s intelligent and human as a real person, then we can make it and the company it’s “dispatched by” accountable for any wrongdoing, right?

    • Fallibilist@feddit.uk
      link
      fedilink
      arrow-up
      4
      arrow-down
      3
      ·
      1 day ago

      That is how slurs tend to be used: as a tool to dehumanize something we wish to grant no moral value to.

      • Handles@leminal.space
        link
        fedilink
        English
        arrow-up
        6
        arrow-down
        1
        ·
        1 day ago

        Except here you’re not dehumanizing as much as working against anthropomorphization.

        • Fallibilist@feddit.uk
          link
          fedilink
          arrow-up
          2
          arrow-down
          2
          ·
          1 day ago

          You’re not trying to remove a moral status but deny it in the first place. That’s different but the end result of it still seems the same.

          We used to justify the mistreatment of animals with the assumption that they had no subjective experience either. I’m not suggesting AI does but I’m also not claiming it with absolute certainty does not either, and the more we develop those systems, the more likely it seems to me that at some point there’s a good chance that they will.

          The psychological drive behind the urge to call something with a derogatory term is the same in both cases.

          • Handles@leminal.space
            link
            fedilink
            English
            arrow-up
            2
            arrow-down
            1
            ·
            22 hours ago

            The psychological drive behind the urge to call something with a derogatory term is the same in both cases.

            Meh. I find that the rhetorical urge to reduce a dichotomy to absurdly general commonalities is the pseudointellectual’s version of a slur.

  • sickday@fedia.io
    link
    fedilink
    arrow-up
    3
    ·
    1 day ago

    “Craftsmen asks mechanics not to throw their wrenches when they get angry” basically

  • corsicanguppy@lemmy.ca
    link
    fedilink
    English
    arrow-up
    4
    arrow-down
    2
    ·
    1 day ago

    I’ve been mandated to use an LLM - but not too much because token$, but efficiently, and whatever - and I get frustrated often. I will write things like ANSWER MY QUESTION and tell it it has a word budget of 10 words. I know I’m gonna trip some criteria and be locked out.

    It’ll be awesome.