Saltar al contenido principal

Gemma 4 31B (Reasoning)

GoogleGemmaOpen WeightApache 2.0 · Uso Comercial

Descripción

Gemma 4 31B is Google DeepMind's flagship dense multimodal model with 31 billion parameters and a 256K context window. Ranks #3 among open models on Arena AI. Built from the same research as Gemini 3, it features Per-Layer Embeddings, Shared KV Cache, alternating sliding-window and global attention, and variable aspect ratio vision encoding. Achieves an estimated LMArena text score of 1452.

Fecha de lanzamiento
2026-04-02
Parámetros
30.7B
Longitud del contexto
262K
Modalidades
audio, image, text, video

Radar de capacidades

17
general
44
coding
86
reasoning
59
science
90
agents
70
multimodal

Rankings

Dominio#PosiciónPuntuaciónFuente
Capacidad agéntica23
58.0
LS
Ranking de codificación239
60.0
AA
Ranking general248
48.0
AA
Ranking multimodal65
54.0
LS
Ciencia172
63.0
AA

Puntuaciones de benchmarks (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench86.4%Aut.

Healthcare

MedXpertQA61.3%Aut.

Language

MMMLU88.4%Aut.
MMLU-Pro85.2%Aut.

Long Context

MRCR v2 (8-needle)66.4%Aut.

Math

AIME 202689.2%Aut.
MathVision85.6%Aut.

Reasoning

GPQANYU + Cohere + Anthropic (2023)84.3%Aut.
LiveCodeBench v680.0%Aut.
BIG-Bench Extra Hard74.4%Aut.
Humanity's Last Exam26.5%Aut.

Vision

MMMU-Pro76.9%Aut.

Índices de evaluación AA

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
85.7
Ifbench(Google Research (2023))
75.6
Lcr(Artificial Analysis)
69.7
Tau2(Sierra + U Toronto + Vector Institute (2025))
59.9
Scicode(UIUC + Argonne National Lab (2024))
45.5
Terminalbench V2 1
43.4
Coding Index(Artificial Analysis)
43.4
Terminalbench Hard(Stanford × Laude Institute (2026))
36.4
Hle(Center for AI Safety + Scale AI (2025))
23.6
Tau Banking
14.8
Intelligence Index(Artificial Analysis)
14.7
Terminalbench V4 0
0.0

Puntuaciones por categoría LLM Stats

(LLM Stats (zeroeval))
Legal
90
Finance
90
Agents
90
Tool Calling
90
Language
80
Physics
80
Biology
80
Chemistry
80
Math
70
Multimodal
70
Reasoning
70
General
70
Long Context
60
Healthcare
60
Vision
60

Precios

Precio de entradaGratis
Precio de salidaGratis
Precio mixto (3:1)Gratis

Velocidad

Tokens/seg35.0
Retraso del primer token0.97s
Tiempo hasta la respuesta50.51s

Ranking de Precios por Proveedor

Ranking de Precios por Proveedor

21 proveedores

Más barato: DeepInfraMás caro: Opper
ProveedorEntradaSalida
1DeepInfraMás barato
$0
$0
2Novita
$0
$0
3FriendliAI
$0
$0
4Together
$0
$0
5Kilo Gateway
$0.06
$0.33
6OpenRouter
$0.09
$0.34
7NanoGPT
$0.1
$0.45
8DevPass (LLM Gateway)
$0.1
$0.25
9CrofAI
$0.1
$0.3
10Lilac
$0.11
$0.35
11FastRouter
$0.13
$0.38
12OrcaRouter
$0.13
$0.38
13Abacus
$0.14
$0.4
14NovitaAI
$0.14
$0.4
15Vercel AI Gateway
$0.14
$0.4
16Merge Gateway
$0.14
$0.4
17Crusoe
$0.14
$0.4
18Neuralwatt
$0.144
$0.42
19ai&
$0.2
$0.5
20Cortecs
$0.223
$0.39
21Opper
$0.46488
$2.44062

Comparar precios entre diferentes proveedores de API para este modelo.

Fuentes externas