메인 콘텐츠로 건너뛰기

Hermes 3 - Llama-3.1 70B

Nous ResearchLlama오픈 웨이트Apache 2.0 · 상업적 사용 가능

설명

Hermes 3 70B is Nous Research's flagship instruction-following model, fine-tuned for advanced reasoning, creative writing, and complex task completion. It features exceptional instruction adherence and strong performance across multiple domains.

출시일
2024-08-15
파라미터
70.0B
컨텍스트 길이
131K
모달리티
text

능력 레이더

20
general
20
coding
25
reasoning
27
science
24
agents
0
multimodal

랭킹

도메인#순위점수소스
코딩 랭킹427
20.0
AA
종합 랭킹483
24.0
AA
과학470
26.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Biology

GPQANYU + Cohere + Anthropic (2023)66.1%자체 보고

Communication

MT-Bench8.99 / 100자체 보고

Finance

MMLU79.1%자체 보고
TruthfulQA63.3%자체 보고
MMLU-Pro47.2%자체 보고

General

PIQA84.4%자체 보고
ARC-E83.0%자체 보고
IFBench81.2%자체 보고
ARC-C65.5%자체 보고
AGIEval56.2%자체 보고
OpenBookQA49.4%자체 보고

Language

BoolQ88.0%자체 보고
Winogrande83.2%자체 보고
BBH67.8%자체 보고

Math

MATH20.8%자체 보고

Reasoning

HellaSwagAI2 (2019)88.2%자체 보고
MuSR50.7%자체 보고

AA 평가 지수

(Artificial Analysis)
Intelligence Index(Artificial Analysis)
4.8
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.6
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
0.5
Gpqa(NYU + Cohere + Anthropic (2023))
0.4
Scicode(UIUC + Argonne National Lab (2024))
0.2
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.2
Hle(Center for AI Safety + Scale AI (2025))
0.0
Aime(MAA (Mathematical Association of America))
0.0

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Roleplay
9
Communication
9
Creativity
9
Reasoning
1
General
1
Physics
80
Instruction Following
80
Language
70
Biology
70
Chemistry
70
Legal
60
Finance
60
Healthcare
60
Math
50

가격

입력 가격$0.7 / 1M 토큰
출력 가격$0.7 / 1M 토큰
혼합 가격 (3:1)$0.7 / 1M 토큰

속도

토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s

공급자 가격 순위

공급자 가격 순위

3개 공급자

최저가: Nous Research최고가: Kilo Gateway
공급자입력출력
1Nous Research주요
$0.7
$0.7
2OpenRouter
$0.7
$0.7
3Kilo Gateway
$0.7
$0.7

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크