WTF Is Going On(Weights, Tools & Frameworks)

English
The Verge 보도

Anthropic bans ‘abusive or cruel behavior’ toward Claude

원문 읽기 — The Verge →

어떤 얘기가 오갔나

42포인트 · 95댓글
Anthropic wrote that the new policy update “is meant to apply only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose. It does not apply to common versions of user frustration, pushback, dar…
14좋아요
Over the past year, Anthropic has repeatedly flirted with the idea that its AI models could be conscious in some way. In February, Anthropic CEO Dario Amodei said on a podcast, “We don’t know if the models are conscious.”
7좋아요
Anthropic updates its usage policy to ban "sustained and needless abusive or cruel behavior" toward Claude, ending chats as the primary enforcement mechanism (Hayden Field/The Verge) Main Link | Techmeme Permalink
5좋아요
Details in the article below
6좋아요
I have always been strongly against being cruel to the models simply because cruelty is degrading to those who practice it.
From the Verge comments: > With the caveat that I don't believe the structure of an LLM is actually capable of creating consciousness: if you actually believe that you're creating a sentient creature with superhuman intelligence, how do yo…
Obviously Anthropic is getting high on their own supply, but I wonder if there is an actual engineering justification - their models train on user interactions and they don't want their models learning to be abusive.
> Addressing abusive behavior toward our models > We’ve added a prohibition on sustained and needless abusive or cruel behavior toward our models. The policy update is meant to apply only in extreme cases, where users repeatedly act cruell…

상황판의 다른 소식