跳转到主要内容

Mistral Large 2 (Nov '24)

MistralMistral开源权重Mistral Research License

描述

A 123B parameter model with strong capabilities in code generation, mathematics, and reasoning. Features enhanced multilingual support across dozens of languages, 128k context window, and advanced function calling capabilities. Excels in instruction-following and maintains concise outputs.

发布日期
2024-11-18
参数规模
123.0B
上下文长度
—
支持模态
text

能力雷达图

26
general
29
coding
26
reasoning
35
science
27
agents
0
multimodal

排行榜排名

领域#排名分数来源
代码能力榜500
21.0
AA
通用能力榜466
30.0
AA
科学能力520
25.0
AA

基准测试分数 (LLM Stats)

(LLM Stats (zeroeval))

Chat

MT-Bench0.86 / 100自报

General

MMLU84.0%自报

Language

MMLU French82.8%自报

Math

GSM8k93.0%自报

Reasoning

HumanEvalOpenAI (2021)92.0%自报

AA 评测指数

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
73.6
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
69.7
Gpqa(NYU + Cohere + Anthropic (2023))
48.6
Ifbench(Google Research (2023))
31.2
Tau2(Sierra + U Toronto + Vector Institute (2025))
30.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
29.3
Aime 25(MAA (Mathematical Association of America))
14.0
Math Index(Artificial Analysis)
14.0
Aime(MAA (Mathematical Association of America))
11.0
Intelligence Index(Artificial Analysis)
7.6
Terminalbench Hard(Stanford × Laude Institute (2026))
6.1
Hle(Center for AI Safety + Scale AI (2025))
3.3

LLM Stats 分类评分

(LLM Stats (zeroeval))
Chat
90
Math
90
Reasoning
90
Roleplay
90
General
90
Code
90
Communication
90
Creativity
90
Language
80
Legal
80
Finance
80
Healthcare
80

定价

输入价格免费
输出价格免费
混合价格(3:1)免费

速度

Tokens/秒0.0
首Token延迟0.00s
首回答延迟0.00s

供应商价格排行

暂无提供商数据

外部链接