Skip to main content

Qwen3 4B 2507 (Reasoning)

AlibabaQwen
Release Date
2025-08-06
Parameters
—
Context Length
—
Modalities
—

Capability Radar

28
general
64
coding
80
reasoning
48
science
75
agents
0
multimodal

Rankings

Domain#RankScoreSource
Code Ranking379
38.0
AA
General Ranking364
37.0
AA
Science404
37.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

No benchmark data available

AA Evaluation Indices

(Artificial Analysis)
Math Index(Artificial Analysis)
82.7
Aime 25(MAA (Mathematical Association of America))
82.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
74.3
Gpqa(NYU + Cohere + Anthropic (2023))
66.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
64.1
Ifbench(Google Research (2023))
49.8
Lcr(Artificial Analysis)
37.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
25.4
Intelligence Index(Artificial Analysis)
8.8
Hle(Center for AI Safety + Scale AI (2025))
6.2
Terminalbench Hard(Stanford × Laude Institute (2026))
1.5

LLM Stats Category Scores

(LLM Stats (zeroeval))

No category score data available

Pricing

Input PriceFree
Output PriceFree
Blended Price (3:1)Free

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

No provider data available

External Sources