Shieldstral 1.0 (3B)
Descripción
Shieldstral is a 3B open-weight multimodal safety classifier from Mistral. It frames content moderation as policy-adaptive yes/no question answering: plain-language policies are supplied at inference time, and the model returns a calibrated safety score from a single forward pass for text, image, or text+image content. Released under Apache 2.0. Aggregate self-reported average F1: text-safety 0.849, multimodal 0.838 (HF mistralai/Shieldstral-1.0-3B / arXiv 2607.25857); per-bench WildGuard/HarmBench/ToxicChat/BeaverTails/Aegis/VLGuard ids are not in the catalog.
Radar de capacidades
Science se estima a partir de las puntuaciones científicas de LLM Stats o del razonamiento cuando no hay benchmarks científicos dedicados.
Rankings
No hay datos de ranking disponibles
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Safety
Índices de evaluación AA
(Artificial Analysis)No hay datos de evaluación AA disponibles
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Precios
No hay datos de precios disponibles
Velocidad
No hay datos de velocidad disponibles
Ranking de Precios por Proveedor
No hay datos de proveedores disponibles