Mistral Large 2 (Nov '24)
MistralMistralओपन वेटMistral Research License
विवरण
A 123B parameter model with strong capabilities in code generation, mathematics, and reasoning. Features enhanced multilingual support across dozens of languages, 128k context window, and advanced function calling capabilities. Excels in instruction-following and maintains concise outputs.
रिलीज़ तिथि
2024-11-18
पैरामीटर
123.0B
संदर्भ लंबाई
—
मोडैलिटीज़
text
क्षमता रडार
26
general
29
coding
26
reasoning
33
science
27
agents
0
multimodal
रैंकिंग
| डोमेन | #रैंक | स्कोर | स्रोत |
|---|---|---|---|
| कोडिंग रैंकिंग | 455 | 16.0 | AA |
| सामान्य रैंकिंग | 399 | 32.0 | AA |
| विज्ञान | 399 | 33.0 | AA |
बेंचमार्क स्कोर (LLM Stats)
(LLM Stats (zeroeval))Code
HumanEvalOpenAI (2021)
92.0%स्वयं
Communication
MT-Bench
0.86 / 100स्वयं
Finance
MMLU
84.0%स्वयं
MMLU French
82.8%स्वयं
Math
GSM8k
93.0%स्वयं
AA मूल्यांकन सूचकांक
(Artificial Analysis)Math Index(Artificial Analysis)14.0
Intelligence Index(Artificial Analysis)9.0
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.7
Gpqa(NYU + Cohere + Anthropic (2023))0.5
Ifbench(Google Research (2023))0.3
Tau2(Sierra + U Toronto + Vector Institute (2025))0.3
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.3
Scicode(UIUC + Argonne National Lab (2024))0.3
Aime 25(MAA (Mathematical Association of America))0.1
Aime(MAA (Mathematical Association of America))0.1
Terminalbench Hard(Stanford × Laude Institute (2026))0.1
Lcr(Artificial Analysis)0.1
Hle(Center for AI Safety + Scale AI (2025))0.0
LLM Stats श्रेणी स्कोर
(LLM Stats (zeroeval))Math90
Reasoning90
Roleplay90
General90
Code90
Communication90
Creativity90
Language80
Legal80
Finance80
Healthcare80
मूल्य निर्धारण
इनपुट मूल्यमुफ्त
आउटपुट मूल्यमुफ्त
मिश्रित मूल्य (3:1)मुफ्त
गति
टोकन/सेकंड0.0
पहले टोकन में देरी0.00s
पहले उत्तर में देरी0.00s
प्रदाता मूल्य रैंकिंग
कोई प्रदाता डेटा उपलब्ध नहीं