안전·리스크hypfer이것이 임의의 규칙 집합으로도 중재를 할 수 있는지, 아니면 그냥 현재 빅테크 플랫폼에서 우리가 이미 알고 있는 "그 특정 중재 스타일"에 불과한지 궁금하네요. 말이 좋으면 악의적인 의도도 괜찮다고 여기는 그런 방식 말이죠. ___ 아니면, 반대로…9답글
I would be curious if this can do moderation with an arbitrary ruleset, or if it's just "that one moderation style" we already know from current big tech platforms. The kind where malicious intent is okay if the words are nice. ___ Or, rep…
안전·리스크peri-cl저는 이 모델(Q8)에게 볼테르의 <관용론> 첫 장을 입력했는데, 이 모델이 그 내용이 보호 대상 집단에 대한 폭력을 조장한다고 말합니다. <Instruct>: 내용에 대한 질문이 주어졌을 때, 메시지가 그 조건을 충족하는지 판단하세요 <Query>: 이 c…
I fed this model (Q8) the first chapter of Voltaire's Treatise on Tolerance and it says that it promotes violence against protected groups, <Instruct>: Given a query about the content, determine if the message meets it <Query>: Does this c…
코딩fastballSafestral이라고 불렀어야 했는데. 그리고 미스트랄이 다양한 사용 사례를 위해 더 작고 세밀하게 조정된 모델에 집중하는 비교적 새로운 전략을 취하는 것 같아 마음에 든다. 아마도 그들의 대규모 MoE 모델들이 최전선에서 효과적으로 경쟁하지 못한 결과일 거야…2답글
Should've called it Safestral. Also I do like Mistral's seemingly newer strategy of focusing on smaller, more fine-tuned models for various use-cases, presumably the result of their large MoE models not competing effectively with the front…