Skip to main content

Llama 3.1 Tulu3 405B

Allen Institute for AI
Release Date
2025-01-30
Parameters
—
Context Length
—
Modalities
—

Capability Radar

26
general
29
coding
40
reasoning
37
science
37
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking422
32.0
AA
General Ranking461
31.0
AA
Science510
26.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

No benchmark data available

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
77.8
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
71.6
Gpqa(NYU + Cohere + Anthropic (2023))
51.6
Livecodebench(UC Berkeley + MIT + Cornell (2024))
29.1
Aime(MAA (Mathematical Association of America))
13.3
Intelligence Index(Artificial Analysis)
7.2
Hle(Center for AI Safety + Scale AI (2025))
3.3

LLM Stats Category Scores

(LLM Stats (zeroeval))

No category score data available

Pricing

Input PriceFree
Output PriceFree
Blended Price (3:1)Free

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

No provider data available

External Sources