메인 콘텐츠로 건너뛰기

Hermes 3 - Llama-3.1 70B

Nous ResearchLlama오픈 웨이트Apache 2.0 · 상업적 사용 가능

설명

Hermes 3 70B is Nous Research's flagship instruction-following model, fine-tuned for advanced reasoning, creative writing, and complex task completion. It features exceptional instruction adherence and strong performance across multiple domains.

출시일
2024-08-15
파라미터
70.0B
컨텍스트 길이
131K
모달리티
text

능력 레이더

21
general
19
coding
25
reasoning
29
science
23
agents
0
multimodal

랭킹

도메인#순위점수소스
코딩 랭킹506
20.0
AA
종합 랭킹540
24.0
AA
과학571
20.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Chat

MT-Bench8.99 / 100자체 보고

General

MMLU79.1%자체 보고
TruthfulQA63.3%자체 보고

Instruction Following

IFBench81.2%자체 보고

Language

BoolQ88.0%자체 보고
MMLU-Pro47.2%자체 보고

Math

MATH20.8%자체 보고

Reasoning

HellaSwagAI2 (2019)88.2%자체 보고
PIQA84.4%자체 보고
Winogrande83.2%자체 보고
ARC-E83.0%자체 보고
BBH67.8%자체 보고
GPQANYU + Cohere + Anthropic (2023)66.1%자체 보고
ARC-C65.5%자체 보고
AGIEval56.2%자체 보고
MuSR50.7%자체 보고
OpenBookQA49.4%자체 보고

AA 평가 지수

(Artificial Analysis)
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
57.1
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
53.8
Gpqa(NYU + Cohere + Anthropic (2023))
40.1
Livecodebench(UC Berkeley + MIT + Cornell (2024))
18.8
Intelligence Index(Artificial Analysis)
6.0
Hle(Center for AI Safety + Scale AI (2025))
4.0
Aime(MAA (Mathematical Association of America))
2.3

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Chat
9
Roleplay
9
Communication
9
Creativity
9
Reasoning
1
General
1
Instruction Following
80
Physics
80
Language
70
Biology
70
Chemistry
70
Legal
60
Finance
60
Healthcare
60
Math
50

가격

입력 가격$0.7 / 1M 토큰
출력 가격$0.7 / 1M 토큰
혼합 가격 (3:1)$0.7 / 1M 토큰

속도

토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s

공급자 가격 순위

공급자 가격 순위

4개 공급자

최저가: DeepInfra최고가: Kilo Gateway
공급자입력출력
1DeepInfra최저가
$0
$0
2Nous Research주요
$0.7
$0.7
3OpenRouter
$0.7
$0.7
4Kilo Gateway
$0.7
$0.7

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크