Mistral Large 2 (Nov '24)
MistralMistralOpen WeightMistral Research License
Descripción
A 123B parameter model with strong capabilities in code generation, mathematics, and reasoning. Features enhanced multilingual support across dozens of languages, 128k context window, and advanced function calling capabilities. Excels in instruction-following and maintains concise outputs.
Fecha de lanzamiento
2024-11-18
Parámetros
123.0B
Longitud del contexto
—
Modalidades
text
Radar de capacidades
26
general
29
coding
26
reasoning
35
science
27
agents
0
multimodal
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Ranking de codificación | 493 | 21.0 | AA |
| Ranking general | 460 | 30.0 | AA |
| Ciencia | 513 | 25.0 | AA |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Chat
MT-Bench
0.86 / 100Aut.
General
MMLU
84.0%Aut.
Language
MMLU French
82.8%Aut.
Math
GSM8k
93.0%Aut.
Reasoning
HumanEvalOpenAI (2021)
92.0%Aut.
Índices de evaluación AA
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))73.6
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))69.7
Gpqa(NYU + Cohere + Anthropic (2023))48.6
Ifbench(Google Research (2023))31.2
Tau2(Sierra + U Toronto + Vector Institute (2025))30.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))29.3
Aime 25(MAA (Mathematical Association of America))14.0
Math Index(Artificial Analysis)14.0
Aime(MAA (Mathematical Association of America))11.0
Intelligence Index(Artificial Analysis)7.6
Terminalbench Hard(Stanford × Laude Institute (2026))6.1
Hle(Center for AI Safety + Scale AI (2025))3.3
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Chat90
Math90
Reasoning90
Roleplay90
General90
Code90
Communication90
Creativity90
Language80
Legal80
Finance80
Healthcare80
Precios
Precio de entradaGratis
Precio de salidaGratis
Precio mixto (3:1)Gratis
Velocidad
Tokens/seg0.0
Retraso del primer token0.00s
Tiempo hasta la respuesta0.00s
Ranking de Precios por Proveedor
No hay datos de proveedores disponibles