Skip to main content

Gemma 4 31B (Reasoning)

GoogleGemmaOpen WeightApache 2.0 · Commercial OK

Description

Gemma 4 31B is Google DeepMind's flagship dense multimodal model with 31 billion parameters and a 256K context window. Ranks #3 among open models on Arena AI. Built from the same research as Gemini 3, it features Per-Layer Embeddings, Shared KV Cache, alternating sliding-window and global attention, and variable aspect ratio vision encoding. Achieves an estimated LMArena text score of 1452.

Release Date
2026-04-02
Parameters
30.7B
Context Length
262K
Modalities
image, text

Capability Radar

28
general
43
coding
86
reasoning
58
science
90
agents
70
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability25
58.0
LS
Code Ranking150
62.0
AA
General Ranking157
60.0
AA
Science108
67.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench86.4%SR

Biology

GPQANYU + Cohere + Anthropic (2023)84.3%SR

Finance

MMLU-Pro85.2%SR

General

MMMLU88.4%SR
LiveCodeBench v680.0%SR
MMMU-Pro76.9%SR
BIG-Bench Extra Hard74.4%SR
MRCR v2 (8-needle)66.4%SR

Healthcare

MedXpertQA61.3%SR

Math

AIME 202689.2%SR
MathVision85.6%SR
Humanity's Last Exam26.5%SR

AA Evaluation Indices

(Artificial Analysis)
Coding Index(Artificial Analysis)
43.4
Intelligence Index(Artificial Analysis)
29.7
Gpqa(NYU + Cohere + Anthropic (2023))
0.9
Ifbench(Google Research (2023))
0.8
Lcr(Artificial Analysis)
0.7
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.6
Terminalbench V2 1
0.4
Scicode(UIUC + Argonne National Lab (2024))
0.4
Terminalbench Hard(Stanford × Laude Institute (2026))
0.4
Hle(Center for AI Safety + Scale AI (2025))
0.2
Tau Banking
0.1

LLM Stats Category Scores

(LLM Stats (zeroeval))
Legal
90
Finance
90
Agents
90
Tool Calling
90
Language
80
Physics
80
Biology
80
Chemistry
80
Math
70
Multimodal
70
Reasoning
70
General
70
Long Context
60
Healthcare
60
Vision
60

Pricing

Input PriceFree
Output PriceFree
Blended Price (3:1)Free

Speed

Tokens/sec35.5
Time to First Token0.99s
Time to Answer49.95s

Provider Price Ranking

Provider Price Ranking

20 providers

Cheapest: DeepInfraMost Expensive: Cerebras
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2FriendliAI
$0
$0
3Novita
$0
$0
4Together
$0
$0
5Kilo Gateway
$0.08
$0.35
6NanoGPT
$0.1
$0.35
7OpenRouter
$0.1
$0.34
8CrofAI
$0.1
$0.3
9LLM Gateway
$0.102
$0.297
10Lilac
$0.11
$0.35
11FastRouter
$0.13
$0.38
12OrcaRouter
$0.13
$0.38
13Abacus
$0.14
$0.4
14NovitaAI
$0.14
$0.4
15Vercel AI Gateway
$0.14
$0.4
16Merge Gateway
$0.14
$0.4
17Neuralwatt
$0.144
$0.42
18ai&
$0.2
$0.5
19Cortecs
$0.223
$0.39
20Cerebras
$0.99
$1.49

Compare pricing across different API providers for this model.

External Sources