Shieldstral 1.0 (3B)
Mistral AI오픈 웨이트Apache 2.0 · 상업적 사용 가능
설명
Shieldstral is a 3B open-weight multimodal safety classifier from Mistral. It frames content moderation as policy-adaptive yes/no question answering: plain-language policies are supplied at inference time, and the model returns a calibrated safety score from a single forward pass for text, image, or text+image content. Released under Apache 2.0. Aggregate self-reported average F1: text-safety 0.849, multimodal 0.838 (HF mistralai/Shieldstral-1.0-3B / arXiv 2607.25857); per-bench WildGuard/HarmBench/ToxicChat/BeaverTails/Aegis/VLGuard ids are not in the catalog.
출시일
2026-08-04
파라미터
3.0B
컨텍스트 길이
—
모달리티
—
능력 레이더
90
general
0
coding
0
reasoning
0
science추정
0
agents
0
multimodal
전용 과학 벤치마크가 없을 때 Science는 LLM Stats 과학 점수 또는 추론 능력에서 추정합니다.
랭킹
랭킹 데이터가 없습니다
벤치마크 점수 (LLM Stats)
(LLM Stats (zeroeval))Safety
XSTest
94.6%자체 보고
AA 평가 지수
(Artificial Analysis)AA 평가 데이터가 없습니다
LLM Stats 카테고리 점수
(LLM Stats (zeroeval))Safety90
General90
가격
가격 데이터가 없습니다
속도
속도 데이터가 없습니다
공급자 가격 순위
프로바이더 데이터가 없습니다