Shieldstral 1.0 (3B)
Mistral AIオープンウエイトApache 2.0 · 商用利用可
説明
Shieldstral is a 3B open-weight multimodal safety classifier from Mistral. It frames content moderation as policy-adaptive yes/no question answering: plain-language policies are supplied at inference time, and the model returns a calibrated safety score from a single forward pass for text, image, or text+image content. Released under Apache 2.0. Aggregate self-reported average F1: text-safety 0.849, multimodal 0.838 (HF mistralai/Shieldstral-1.0-3B / arXiv 2607.25857); per-bench WildGuard/HarmBench/ToxicChat/BeaverTails/Aegis/VLGuard ids are not in the catalog.
リリース日
2026-08-04
パラメータ
3.0B
コンテキスト長
—
モダリティ
—
能力レーダー
90
general
0
coding
0
reasoning
0
science推定
0
agents
0
multimodal
専用の科学ベンチマークがない場合、Science は LLM Stats の科学スコアまたは推論能力から推定します。
ランキング
ランキングデータがありません
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Safety
XSTest
94.6%自己申告
AA評価指数
(Artificial Analysis)AA評価データがありません
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Safety90
General90
価格設定
価格データがありません
速度
速度データがありません
プロバイダー価格ランキング
プロバイダーデータがありません