Llama 3.1 Nemotron Instruct 70B
NVIDIALlamaオープンウエイトLlama 3.1 Community License
説明
A large language model customized by NVIDIA to improve the helpfulness of LLM generated responses. It is a fine-tuned version of Llama 3.1 70B Instruct. The model was trained using RLHF (REINFORCE) with HelpSteer2-Preference prompts.
リリース日
2024-10-15
パラメータ
70.0B
コンテキスト長
128K
モダリティ
text
能力レーダー
25
general
17
coding
27
reasoning
33
science
24
agents
0
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 576 | 11.0 | AA |
| 総合ランキング | 502 | 28.0 | AA |
| 科学 | 530 | 24.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Chat
MT-Bench
0.09 / 100自己申告
General
MMLU
80.2%自己申告
Instruct HumanEval
73.8%自己申告
TruthfulQA
58.6%自己申告
Language
MMLU Chat
80.6%自己申告
Math
GSM8k
91.4%自己申告
GSM8K Chat
81.9%自己申告
Reasoning
HellaSwagAI2 (2019)
85.6%自己申告
Winogrande
84.5%自己申告
ARC-C
69.2%自己申告
Summarization
XLSum English
31.6%自己申告
AA評価指数
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))73.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))69.0
Gpqa(NYU + Cohere + Anthropic (2023))46.5
Ifbench(Google Research (2023))30.7
Aime(MAA (Mathematical Association of America))24.7
Tau2(Sierra + U Toronto + Vector Institute (2025))23.1
Livecodebench(UC Berkeley + MIT + Cornell (2024))16.9
Math Index(Artificial Analysis)11.0
Aime 25(MAA (Mathematical Association of America))11.0
Lcr(Artificial Analysis)8.3
Intelligence Index(Artificial Analysis)6.9
Terminalbench Hard(Stanford × Laude Institute (2026))4.5
Hle(Center for AI Safety + Scale AI (2025))4.2
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Math90
Language80
Legal70
Reasoning70
Finance70
General70
Healthcare70
Chat10
Roleplay10
Communication10
Creativity10
価格設定
入力価格無料
出力価格無料
混合価格(3:1)無料
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダーデータがありません