跳轉到主要內容

Shieldstral 1.0 (3B)

Mistral AI開源權重Apache 2.0 · 商用許可

描述

Shieldstral is a 3B open-weight multimodal safety classifier from Mistral. It frames content moderation as policy-adaptive yes/no question answering: plain-language policies are supplied at inference time, and the model returns a calibrated safety score from a single forward pass for text, image, or text+image content. Released under Apache 2.0. Aggregate self-reported average F1: text-safety 0.849, multimodal 0.838 (HF mistralai/Shieldstral-1.0-3B / arXiv 2607.25857); per-bench WildGuard/HarmBench/ToxicChat/BeaverTails/Aegis/VLGuard ids are not in the catalog.

發布日期
2026-08-04
參數規模
3.0B
上下文長度
支援模態

能力雷達圖

90
general
0
coding
0
reasoning
0
science估算
0
agents
0
multimodal

缺少專門科學評測時,Science 由 LLM Stats 科學得分或推理能力估算。

排行榜排名

暫無排名資料

基準測試分數 (LLM Stats)

(LLM Stats (zeroeval))

Safety

XSTest94.6%自報

AA 評測指數

(Artificial Analysis)

暫無 AA 評測資料

LLM Stats 分類評分

(LLM Stats (zeroeval))
Safety
90
General
90

定價

暫無定價資料

速度

暫無速度資料

供應商價格排行

暫無提供商資料

外部連結