3B 오픈 가중치 모델 'Shieldstral'이 멀티모달 콘텐츠 검열을 위한 안전 분류기로 공개되었다.
'Shieldstral'은 3B 멀티모달 안전 분류기로, 제품별로 적용할 수 있는 안전 정책을 추론할 수 있습니다. 이 모델은 고정된 유해성 분류 체계 대신 자연어 정책을 이용하여 텍스트와 이미지를 평가합니다. Apache 2.0 라이센스로 공개되었습니다.
'Shieldstral', a 3B open-weight model for multimodal content censorship, has been released.
'Shieldstral' is a 3B multimodal safety classifier that can infer safety policies applicable per product. Instead of a fixed toxicity classification system, it utilizes natural language policies to evaluate text and images. It has been released under the Apache 2.0 license.