Skip to main content

Llama 4 Maverick

MetaLlamaOpen WeightLlama 4 Community License Agreement

Description

Llama 4 Maverick is a natively multimodal model capable of processing both text and images. It features a 17 billion active parameter mixture-of-experts (MoE) architecture with 128 experts, supporting a wide range of multimodal tasks such as conversational interaction, image analysis, and code generation. The model includes a 1 million token context window.

Release Date
2025-04-05
Parameters
400.0B
Context Length
1.0M
Modalities
image, text

Capability Radar

31
general
26
coding
39
reasoning
42
science
35
agents
90
multimodal

Rankings

Domain#RankScoreSource
Code Ranking433
30.0
AA
General Ranking387
35.0
AA
Multimodal Ranking104
45.0
LS
Science393
38.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

General

MMLU85.5%SR

Language

MMLU-Pro80.5%SR
TydiQA31.7%SR

Math

MGSM92.3%SR
MathVista73.7%SR
MATH61.2%SR

Multimodal

MMMU73.4%SR

Reasoning

ChartQAMasry et al. (2022)90.0%SR
MBPP0.78 / 100SR
GPQANYU + Cohere + Anthropic (2023)69.8%SR
LiveCodeBench43.4%SR

Vision

DocVQADocVQA (2020)94.4%SR
MMMU-Pro59.6%SR

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
88.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
80.9
Gpqa(NYU + Cohere + Anthropic (2023))
67.1
Lcr(Artificial Analysis)
50.0
Ifbench(Google Research (2023))
43.0
Livecodebench(UC Berkeley + MIT + Cornell (2024))
39.7
Aime(MAA (Mathematical Association of America))
39.0
Scicode(UIUC + Argonne National Lab (2024))
31.7
Aime 25(MAA (Mathematical Association of America))
19.3
Math Index(Artificial Analysis)
19.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
17.8
Coding Index(Artificial Analysis)
16.3
Intelligence Index(Artificial Analysis)
10.0
Terminalbench V2 1
7.9
Terminalbench Hard(Stanford × Laude Institute (2026))
6.8
Hle(Center for AI Safety + Scale AI (2025))
4.9
Tau Banking
3.7
Terminalbench V4 0
0.0

LLM Stats Category Scores

(LLM Stats (zeroeval))
Image To Text
90
Language
80
Legal
80
Math
80
Multimodal
80
Finance
80
General
80
Healthcare
80
Vision
80
Physics
70
Reasoning
70
Biology
70
Chemistry
70
Code
40

Pricing

Input Price$0.26 / 1M tokens
Output Price$0.91 / 1M tokens
Blended Price (3:1)$0.422 / 1M tokens

Speed

Tokens/sec86.2
Time to First Token0.59s
Time to Answer0.59s

Provider Price Ranking

Provider Price Ranking

8 providers

Cheapest: DeepInfraMost Expensive: Neon
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2NanoGPT
$0.15
$0.6
3Helicone
$0.15
$0.6
4OpenRouter
$0.1875
$0.6525
5Kilo Gateway
$0.1875
$0.6525
6DigitalOcean
$0.25
$0.87
7MetaPRIMARY
$0.26
$0.91
8Neon
$0.5
$1.5

Compare pricing across different API providers for this model.

External Sources