Shieldstral 1.0 (3B)
Mistral AI开源权重Apache 2.0 · 商用许可
描述
Shieldstral is a 3B open-weight multimodal safety classifier from Mistral. It frames content moderation as policy-adaptive yes/no question answering: plain-language policies are supplied at inference time, and the model returns a calibrated safety score from a single forward pass for text, image, or text+image content. Released under Apache 2.0. Aggregate self-reported average F1: text-safety 0.849, multimodal 0.838 (HF mistralai/Shieldstral-1.0-3B / arXiv 2607.25857); per-bench WildGuard/HarmBench/ToxicChat/BeaverTails/Aegis/VLGuard ids are not in the catalog.
发布日期
2026-08-04
参数规模
3.0B
上下文长度
—
支持模态
—
能力雷达图
90
general
0
coding
0
reasoning
0
science估算
0
agents
0
multimodal
缺少专门科学评测时,Science 由 LLM Stats 科学得分或推理能力估算。
排行榜排名
暂无排名数据
基准测试分数 (LLM Stats)
(LLM Stats (zeroeval))Safety
XSTest
94.6%自报
AA 评测指数
(Artificial Analysis)暂无 AA 评测数据
LLM Stats 分类评分
(LLM Stats (zeroeval))Safety90
General90
定价
暂无定价数据
速度
暂无速度数据
供应商价格排行
暂无提供商数据