Mistral Large 2 (Nov '24)
MistralMistralオープンウエイトMistral Research License
説明
A 123B parameter model with strong capabilities in code generation, mathematics, and reasoning. Features enhanced multilingual support across dozens of languages, 128k context window, and advanced function calling capabilities. Excels in instruction-following and maintains concise outputs.
リリース日
2024-11-18
パラメータ
123.0B
コンテキスト長
—
モダリティ
text
能力レーダー
26
general
29
coding
26
reasoning
35
science
27
agents
0
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 493 | 21.0 | AA |
| 総合ランキング | 460 | 30.0 | AA |
| 科学 | 513 | 25.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Chat
MT-Bench
0.86 / 100自己申告
General
MMLU
84.0%自己申告
Language
MMLU French
82.8%自己申告
Math
GSM8k
93.0%自己申告
Reasoning
HumanEvalOpenAI (2021)
92.0%自己申告
AA評価指数
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))73.6
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))69.7
Gpqa(NYU + Cohere + Anthropic (2023))48.6
Ifbench(Google Research (2023))31.2
Tau2(Sierra + U Toronto + Vector Institute (2025))30.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))29.3
Aime 25(MAA (Mathematical Association of America))14.0
Math Index(Artificial Analysis)14.0
Aime(MAA (Mathematical Association of America))11.0
Intelligence Index(Artificial Analysis)7.6
Terminalbench Hard(Stanford × Laude Institute (2026))6.1
Hle(Center for AI Safety + Scale AI (2025))3.3
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Chat90
Math90
Reasoning90
Roleplay90
General90
Code90
Communication90
Creativity90
Language80
Legal80
Finance80
Healthcare80
価格設定
入力価格無料
出力価格無料
混合価格(3:1)無料
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダーデータがありません