Gemma 4 31B (Reasoning)
GoogleGemmaOpen WeightApache 2.0 · Commercial OK
Description
Gemma 4 31B is Google DeepMind's flagship dense multimodal model with 31 billion parameters and a 256K context window. Ranks #3 among open models on Arena AI. Built from the same research as Gemini 3, it features Per-Layer Embeddings, Shared KV Cache, alternating sliding-window and global attention, and variable aspect ratio vision encoding. Achieves an estimated LMArena text score of 1452.
Release Date
2026-04-02
Parameters
30.7B
Context Length
262K
Modalities
image, text
Capability Radar
28
general
43
coding
86
reasoning
58
science
90
agents
70
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 25 | 58.0 | LS |
| Code Ranking | 150 | 62.0 | AA |
| General Ranking | 157 | 60.0 | AA |
| Science | 108 | 67.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
86.4%SR
Biology
GPQANYU + Cohere + Anthropic (2023)
84.3%SR
Finance
MMLU-Pro
85.2%SR
General
MMMLU
88.4%SR
LiveCodeBench v6
80.0%SR
MMMU-Pro
76.9%SR
BIG-Bench Extra Hard
74.4%SR
MRCR v2 (8-needle)
66.4%SR
Healthcare
MedXpertQA
61.3%SR
Math
AIME 2026
89.2%SR
MathVision
85.6%SR
Humanity's Last Exam
26.5%SR
AA Evaluation Indices
(Artificial Analysis)Coding Index(Artificial Analysis)43.4
Intelligence Index(Artificial Analysis)29.7
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Ifbench(Google Research (2023))0.8
Lcr(Artificial Analysis)0.7
Tau2(Sierra + U Toronto + Vector Institute (2025))0.6
Terminalbench V2 10.4
Scicode(UIUC + Argonne National Lab (2024))0.4
Terminalbench Hard(Stanford × Laude Institute (2026))0.4
Hle(Center for AI Safety + Scale AI (2025))0.2
Tau Banking0.1
LLM Stats Category Scores
(LLM Stats (zeroeval))Legal90
Finance90
Agents90
Tool Calling90
Language80
Physics80
Biology80
Chemistry80
Math70
Multimodal70
Reasoning70
General70
Long Context60
Healthcare60
Vision60
Pricing
Input PriceFree
Output PriceFree
Blended Price (3:1)Free
Speed
Tokens/sec35.5
Time to First Token0.99s
Time to Answer49.95s
Provider Price Ranking
Provider Price Ranking
20 providers
Cheapest: DeepInfraMost Expensive: Cerebras
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2FriendliAI
$0
$0
3Novita
$0
$0
4Together
$0
$0
5Kilo Gateway
$0.08
$0.35
6NanoGPT
$0.1
$0.35
7OpenRouter
$0.1
$0.34
8CrofAI
$0.1
$0.3
9LLM Gateway
$0.102
$0.297
10Lilac
$0.11
$0.35
11FastRouter
$0.13
$0.38
12OrcaRouter
$0.13
$0.38
13Abacus
$0.14
$0.4
14NovitaAI
$0.14
$0.4
15Vercel AI Gateway
$0.14
$0.4
16Merge Gateway
$0.14
$0.4
17Neuralwatt
$0.144
$0.42
18ai&
$0.2
$0.5
19Cortecs
$0.223
$0.39
20Cerebras
$0.99
$1.49
Compare pricing across different API providers for this model.