跳轉到主要內容

Mistral Large 2 (Nov '24)

MistralMistral開源權重Mistral Research License

描述

A 123B parameter model with strong capabilities in code generation, mathematics, and reasoning. Features enhanced multilingual support across dozens of languages, 128k context window, and advanced function calling capabilities. Excels in instruction-following and maintains concise outputs.

發布日期
2024-11-18
參數規模
123.0B
上下文長度
—
支援模態
text

能力雷達圖

26
general
29
coding
26
reasoning
35
science
27
agents
0
multimodal

排行榜排名

領域#排名分數來源
程式碼能力榜493
21.0
AA
通用能力榜460
30.0
AA
科學能力513
25.0
AA

基準測試分數 (LLM Stats)

(LLM Stats (zeroeval))

Chat

MT-Bench0.86 / 100自報

General

MMLU84.0%自報

Language

MMLU French82.8%自報

Math

GSM8k93.0%自報

Reasoning

HumanEvalOpenAI (2021)92.0%自報

AA 評測指數

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
73.6
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
69.7
Gpqa(NYU + Cohere + Anthropic (2023))
48.6
Ifbench(Google Research (2023))
31.2
Tau2(Sierra + U Toronto + Vector Institute (2025))
30.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
29.3
Aime 25(MAA (Mathematical Association of America))
14.0
Math Index(Artificial Analysis)
14.0
Aime(MAA (Mathematical Association of America))
11.0
Intelligence Index(Artificial Analysis)
7.6
Terminalbench Hard(Stanford × Laude Institute (2026))
6.1
Hle(Center for AI Safety + Scale AI (2025))
3.3

LLM Stats 分類評分

(LLM Stats (zeroeval))
Chat
90
Math
90
Reasoning
90
Roleplay
90
General
90
Code
90
Communication
90
Creativity
90
Language
80
Legal
80
Finance
80
Healthcare
80

定價

輸入價格免費
輸出價格免費
混合價格(3:1)免費

速度

Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s

供應商價格排行

暫無提供商資料

外部連結