Mistral Large 2 (Nov '24)
MistralMistralОткрытые весаMistral Research License
Описание
A 123B parameter model with strong capabilities in code generation, mathematics, and reasoning. Features enhanced multilingual support across dozens of languages, 128k context window, and advanced function calling capabilities. Excels in instruction-following and maintains concise outputs.
Дата выхода
2024-11-18
Параметры
123.0B
Длина контекста
—
Модальности
text
Радар способностей
26
general
29
coding
26
reasoning
35
science
27
agents
0
multimodal
Рейтинги
| Домен | #Место | Оценка | Источник |
|---|---|---|---|
| Рейтинг кодинга | 493 | 21.0 | AA |
| Общий рейтинг | 460 | 30.0 | AA |
| Наука | 513 | 25.0 | AA |
Оценки бенчмарков (LLM Stats)
(LLM Stats (zeroeval))Chat
MT-Bench
0.86 / 100Сам.
General
MMLU
84.0%Сам.
Language
MMLU French
82.8%Сам.
Math
GSM8k
93.0%Сам.
Reasoning
HumanEvalOpenAI (2021)
92.0%Сам.
Индексы оценки AA
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))73.6
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))69.7
Gpqa(NYU + Cohere + Anthropic (2023))48.6
Ifbench(Google Research (2023))31.2
Tau2(Sierra + U Toronto + Vector Institute (2025))30.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))29.3
Aime 25(MAA (Mathematical Association of America))14.0
Math Index(Artificial Analysis)14.0
Aime(MAA (Mathematical Association of America))11.0
Intelligence Index(Artificial Analysis)7.6
Terminalbench Hard(Stanford × Laude Institute (2026))6.1
Hle(Center for AI Safety + Scale AI (2025))3.3
Оценки категорий LLM Stats
(LLM Stats (zeroeval))Chat90
Math90
Reasoning90
Roleplay90
General90
Code90
Communication90
Creativity90
Language80
Legal80
Finance80
Healthcare80
Цены
Цена вводаБесплатно
Цена выводаБесплатно
Смешанная цена (3:1)Бесплатно
Скорость
Токенов/сек0.0
Задержка первого токена0.00s
Время до первого ответа0.00s
Рейтинг цен провайдеров
Нет данных провайдеров