Llama 3.3 Instruct 70B
MetaLlamaOpen WeightLlama 3.3 Community License Agreement
Description
Llama 3.3 is a multilingual large language model optimized for dialogue use cases across multiple languages. It is a pretrained and instruction-tuned generative model with 70 billion parameters, outperforming many open-source and closed chat models on common industry benchmarks. Llama 3.3 supports a context length of 128,000 tokens and is designed for commercial and research use in multiple languages.
Release Date
2024-12-06
Parameters
70.0B
Context Length
131K
Modalities
text
Capability Radar
26
general
18
coding
28
reasoning
36
science
80
agents
0
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 532 | 17.0 | AA |
| General Ranking | 409 | 34.0 | AA |
| Science | 519 | 25.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Chat
IFEvalGoogle Research (2023)
92.1%SR
General
MMLU
86.0%SR
BFCL v2
77.3%SR
Language
MMLU-Pro
68.9%SR
Math
MGSM
91.1%SR
MATH
77.0%SR
Reasoning
HumanEvalOpenAI (2021)
88.4%SR
MBPP EvalPlus
87.6%SR
GPQANYU + Cohere + Anthropic (2023)
50.5%SR
AA Evaluation Indices
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))77.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))71.3
Gpqa(NYU + Cohere + Anthropic (2023))49.8
Ifbench(Google Research (2023))47.1
Aime(MAA (Mathematical Association of America))30.0
Livecodebench(UC Berkeley + MIT + Cornell (2024))28.8
Tau2(Sierra + U Toronto + Vector Institute (2025))26.6
Lcr(Artificial Analysis)15.7
Coding Index(Artificial Analysis)11.9
Intelligence Index(Artificial Analysis)7.7
Math Index(Artificial Analysis)7.7
Aime 25(MAA (Mathematical Association of America))7.7
Terminalbench V2 14.9
Hle(Center for AI Safety + Scale AI (2025))3.6
Terminalbench Hard(Stanford × Laude Institute (2026))3.0
LLM Stats Category Scores
(LLM Stats (zeroeval))Chat90
Instruction Following90
Structured Output90
Code90
Language80
Legal80
Math80
Reasoning80
Finance80
General80
Healthcare80
Tool Calling80
Physics50
Biology50
Chemistry50
Pricing
Input Price$0.71 / 1M tokens
Output Price$0.72 / 1M tokens
Blended Price (3:1)$0.712 / 1M tokens
Speed
Tokens/sec90.7
Time to First Token0.64s
Time to Answer0.64s
Provider Price Ranking
Provider Price Ranking
6 providers
Cheapest: DeepInfraMost Expensive: Meta
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2NanoGPT
$0.05
$0.23
3OpenRouter
$0.1
$0.32
4Kilo Gateway
$0.1
$0.32
5NovitaAI
$0.135
$0.4
6MetaPRIMARY
$0.71
$0.72
Compare pricing across different API providers for this model.