Anthropic wrote that the new policy update “is meant to apply only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose. It does not apply to common versions of user frustration, pushback, dar…
Over the past year, Anthropic has repeatedly flirted with the idea that its AI models could be conscious in some way. In February, Anthropic CEO Dario Amodei said on a podcast, “We don’t know if the models are conscious.”
Anthropic updates its usage policy to ban "sustained and needless abusive or cruel behavior" toward Claude, ending chats as the primary enforcement mechanism (Hayden Field/The Verge) Main Link | Techmeme Permalink
From the Verge comments: > With the caveat that I don't believe the structure of an LLM is actually capable of creating consciousness: if you actually believe that you're creating a sentient creature with superhuman intelligence, how do yo…
Obviously Anthropic is getting high on their own supply, but I wonder if there is an actual engineering justification - their models train on user interactions and they don't want their models learning to be abusive.
> Addressing abusive behavior toward our models > We’ve added a prohibition on sustained and needless abusive or cruel behavior toward our models. The policy update is meant to apply only in extreme cases, where users repeatedly act cruell…