메인 콘텐츠로 건너뛰기

Llama 3.1 Nemotron Instruct 70B

NVIDIALlama오픈 웨이트Llama 3.1 Community License

설명

A large language model customized by NVIDIA to improve the helpfulness of LLM generated responses. It is a fine-tuned version of Llama 3.1 70B Instruct. The model was trained using RLHF (REINFORCE) with HelpSteer2-Preference prompts.

출시일
2024-10-15
파라미터
70.0B
컨텍스트 길이
128K
모달리티
text

능력 레이더

25
general
18
coding
27
reasoning
30
science
24
agents
0
multimodal

랭킹

도메인#순위점수소스
코딩 랭킹498
11.0
AA
종합 랭킹433
29.0
AA
과학440
29.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Communication

MT-Bench0.09 / 100자체 보고

Finance

MMLU Chat80.6%자체 보고
MMLU80.2%자체 보고
TruthfulQA58.6%자체 보고

General

Instruct HumanEval73.8%자체 보고
ARC-C69.2%자체 보고

Language

Winogrande84.5%자체 보고
XLSum English31.6%자체 보고

Math

GSM8k91.4%자체 보고
GSM8K Chat81.9%자체 보고

Reasoning

HellaSwagAI2 (2019)85.6%자체 보고

AA 평가 지수

(Artificial Analysis)
Math Index(Artificial Analysis)
11.0
Intelligence Index(Artificial Analysis)
7.4
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
0.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.7
Gpqa(NYU + Cohere + Anthropic (2023))
0.5
Ifbench(Google Research (2023))
0.3
Aime(MAA (Mathematical Association of America))
0.2
Scicode(UIUC + Argonne National Lab (2024))
0.2
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.2
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.2
Aime 25(MAA (Mathematical Association of America))
0.1
Lcr(Artificial Analysis)
0.1
Terminalbench Hard(Stanford × Laude Institute (2026))
0.0
Hle(Center for AI Safety + Scale AI (2025))
0.0

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Math
90
Language
80
Legal
70
Reasoning
70
Finance
70
General
70
Healthcare
70
Roleplay
10
Communication
10
Creativity
10

가격

입력 가격$1.2 / 1M 토큰
출력 가격$1.2 / 1M 토큰
혼합 가격 (3:1)$1.2 / 1M 토큰

속도

토큰/초82.3
첫 토큰 지연5.52s
첫 응답 지연5.52s

공급자 가격 순위

공급자 가격 순위

1개 공급자

공급자입력출력
1NVIDIA주요
$1.2
$1.2

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크