跳轉到主要內容

ERNIE 5.0 Thinking Preview

BaiduProprietary

描述

ERNIE 5.0 is Baidu's flagship foundation model featuring state-of-the-art performance across reasoning, mathematics, coding, and knowledge benchmarks. It represents a significant advancement over ERNIE 4.5 with enhanced multilingual capabilities, improved instruction following, and stronger agentic task performance.

發布日期
2025-11-13
參數規模
上下文長度
支援模態

能力雷達圖

39
general
71
coding
84
reasoning
51
science
80
agents
40
multimodal

排行榜排名

領域#排名分數來源
程式碼能力榜252
45.0
AA
通用能力榜183
57.0
AA
科學能力190
55.0
AA

基準測試分數 (LLM Stats)

(LLM Stats (zeroeval))

Biology

GPQANYU + Cohere + Anthropic (2023)85.0%自報

Factuality

SimpleQA75.0%自報

Finance

MMLU-Pro87.0%自報

Math

AIME 202587.0%自報
Humanity's Last Exam39.0%自報

AA 評測指數

(Artificial Analysis)
Math Index(Artificial Analysis)
85.0
Intelligence Index(Artificial Analysis)
22.3
Aime 25(MAA (Mathematical Association of America))
0.8
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.8
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.8
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.8
Ifbench(Google Research (2023))
0.4
Scicode(UIUC + Argonne National Lab (2024))
0.4
Terminalbench Hard(Stanford × Laude Institute (2026))
0.3
Hle(Center for AI Safety + Scale AI (2025))
0.1
Lcr(Artificial Analysis)
0.1

LLM Stats 分類評分

(LLM Stats (zeroeval))
Legal
90
Language
90
Finance
90
Healthcare
90
Physics
80
Factuality
80
Biology
80
Chemistry
80
Math
70
Reasoning
70
General
70
Vision
40

定價

輸入價格免費
輸出價格免費
混合價格(3:1)免費

速度

Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s

供應商價格排行

暫無提供商資料

外部連結