Mistral Large 2 (Nov '24)
MistralMistralОткрытые весаMistral Research License
Описание
A 123B parameter model with strong capabilities in code generation, mathematics, and reasoning. Features enhanced multilingual support across dozens of languages, 128k context window, and advanced function calling capabilities. Excels in instruction-following and maintains concise outputs.
Дата выхода
2024-11-18
Параметры
123.0B
Длина контекста
—
Модальности
text
Радар способностей
26
general
29
coding
26
reasoning
33
science
27
agents
0
multimodal
Рейтинги
| Домен | #Место | Оценка | Источник |
|---|---|---|---|
| Рейтинг кодинга | 459 | 16.0 | AA |
| Общий рейтинг | 403 | 32.0 | AA |
| Наука | 403 | 33.0 | AA |
Оценки бенчмарков (LLM Stats)
(LLM Stats (zeroeval))Code
HumanEvalOpenAI (2021)
92.0%Сам.
Communication
MT-Bench
0.86 / 100Сам.
Finance
MMLU
84.0%Сам.
MMLU French
82.8%Сам.
Math
GSM8k
93.0%Сам.
Индексы оценки AA
(Artificial Analysis)Math Index(Artificial Analysis)14.0
Intelligence Index(Artificial Analysis)9.0
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.7
Gpqa(NYU + Cohere + Anthropic (2023))0.5
Ifbench(Google Research (2023))0.3
Tau2(Sierra + U Toronto + Vector Institute (2025))0.3
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.3
Scicode(UIUC + Argonne National Lab (2024))0.3
Aime 25(MAA (Mathematical Association of America))0.1
Aime(MAA (Mathematical Association of America))0.1
Terminalbench Hard(Stanford × Laude Institute (2026))0.1
Lcr(Artificial Analysis)0.1
Hle(Center for AI Safety + Scale AI (2025))0.0
Оценки категорий LLM Stats
(LLM Stats (zeroeval))Math90
Reasoning90
Roleplay90
General90
Code90
Communication90
Creativity90
Legal80
Language80
Finance80
Healthcare80
Цены
Цена вводаБесплатно
Цена выводаБесплатно
Смешанная цена (3:1)Бесплатно
Скорость
Токенов/сек0.0
Задержка первого токена0.00s
Время до первого ответа0.00s
Рейтинг цен провайдеров
Нет данных провайдеров