Skip to main content

Llama 3.1 Instruct 70B

MetaLlamaOpen WeightLlama 3.1 Community License

Description

Llama 3.1 70B Instruct is a large language model optimized for multilingual dialogue use cases. It outperforms many available open source and closed chat models on common industry benchmarks.

Release Date
2024-07-23
Parameters
70.0B
Context Length
131K
Modalities
text

Capability Radar

25
general
23
coding
20
reasoning
30
science
70
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking547
15.0
AA
General Ranking516
27.0
AA
Science564
21.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Chat

IFEvalGoogle Research (2023)87.5%SR

General

BFCL84.8%SR
MMLU83.6%SR
Multipl-E MBPP62.0%SR
Nexus56.7%SR

Language

MMLU (CoT)86.0%SR
MMLU-Pro66.4%SR
Multipl-E HumanEval65.5%SR

Math

GSM-8K (CoT)95.1%SR
Multilingual MGSM (CoT)86.9%SR
MATH (CoT)68.0%SR

Reasoning

ARC-C94.8%SR
API-Bank90.0%SR
MBPP ++ base version86.0%SR
HumanEvalOpenAI (2021)80.5%SR
DROP79.6%SR
GPQANYU + Cohere + Anthropic (2023)41.7%SR
Gorilla Benchmark API Bench29.7%SR

AA Evaluation Indices

(Artificial Analysis)
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
67.6
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
64.9
Gpqa(NYU + Cohere + Anthropic (2023))
40.9
Ifbench(Google Research (2023))
34.4
Livecodebench(UC Berkeley + MIT + Cornell (2024))
23.2
Aime(MAA (Mathematical Association of America))
17.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
15.2
Intelligence Index(Artificial Analysis)
6.6
Hle(Center for AI Safety + Scale AI (2025))
4.5
Math Index(Artificial Analysis)
4.0
Aime 25(MAA (Mathematical Association of America))
4.0
Terminalbench Hard(Stanford × Laude Institute (2026))
3.0

LLM Stats Category Scores

(LLM Stats (zeroeval))
Chat
90
Instruction Following
90
Structured Output
90
Language
80
Legal
80
Math
80
Finance
80
Healthcare
80
Reasoning
70
General
70
Tool Calling
70
Code
60
Physics
40
Biology
40
Chemistry
40

Pricing

Input Price$0.56 / 1M tokens
Output Price$0.56 / 1M tokens
Blended Price (3:1)$0.56 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

4 providers

Cheapest: DeepInfraMost Expensive: Meta
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2OpenRouter
$0.4
$0.4
3Kilo Gateway
$0.4
$0.4
4MetaPRIMARY
$0.56
$0.56

Compare pricing across different API providers for this model.

External Sources