WTF Is Going On(Weights, Tools & Frameworks)

한국어
The Verge filed

Anthropic bans ‘abusive or cruel behavior’ toward Claude

read it at the source — The Verge →

What was said about it

42 points · 95 comments
Anthropic wrote that the new policy update “is meant to apply only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose. It does not apply to common versions of user frustration, pushback, dar…
14 likes
Over the past year, Anthropic has repeatedly flirted with the idea that its AI models could be conscious in some way. In February, Anthropic CEO Dario Amodei said on a podcast, “We don’t know if the models are conscious.”
7 likes
Anthropic updates its usage policy to ban "sustained and needless abusive or cruel behavior" toward Claude, ending chats as the primary enforcement mechanism (Hayden Field/The Verge) Main Link | Techmeme Permalink
5 likes
Details in the article below
6 likes
I have always been strongly against being cruel to the models simply because cruelty is degrading to those who practice it.
From the Verge comments: > With the caveat that I don't believe the structure of an LLM is actually capable of creating consciousness: if you actually believe that you're creating a sentient creature with superhuman intelligence, how do yo…
Obviously Anthropic is getting high on their own supply, but I wonder if there is an actual engineering justification - their models train on user interactions and they don't want their models learning to be abusive.
> Addressing abusive behavior toward our models > We’ve added a prohibition on sustained and needless abusive or cruel behavior toward our models. The policy update is meant to apply only in extreme cases, where users repeatedly act cruell…

Also on the board