Saltar al contenido principal

Gemma 3 12B Instruct

GoogleGemmaOpen WeightGemma · Uso Comercial

Descripción

Gemma 3 12B is a 12-billion-parameter vision-language model from Google, handling text and image input and generating text output. It features a 128K context window, multilingual support, and open weights. Suitable for question answering, summarization, reasoning, and image understanding tasks.

Fecha de lanzamiento
2025-03-12
Parámetros
12.0B
Longitud del contexto
131K
Modalidades
image, text

Radar de capacidades

22
general
10
coding
31
reasoning
23
science
25
agents
80
multimodal

Rankings

Dominio#PosiciónPuntuaciónFuente
Ranking de codificación519
8.0
AA
Ranking general474
25.0
AA
Ranking multimodal57
43.0
LS
Ciencia501
21.0
AA

Puntuaciones de benchmarks (LLM Stats)

(LLM Stats (zeroeval))

Biology

GPQANYU + Cohere + Anthropic (2023)40.9%Aut.

Code

HumanEvalOpenAI (2021)85.4%Aut.
LiveCodeBench24.6%Aut.

Factuality

FACTS Grounding75.8%Aut.
SimpleQA6.3%Aut.

Finance

MMLU-Pro60.6%Aut.

General

IFEvalGoogle Research (2023)88.9%Aut.
Natural2Code80.7%Aut.
MBPP0.73 / 100Aut.
Global-MMLU-Lite69.5%Aut.
MMMU (val)59.6%Aut.
BIG-Bench Extra Hard16.3%Aut.

Image To Text

DocVQADocVQA (2020)87.1%Aut.
VQAv2 (val)71.6%Aut.
TextVQA67.7%Aut.

Language

BIG-Bench Hard85.7%Aut.
WMT24++51.6%Aut.
ECLeKTic10.3%Aut.

Math

GSM8k94.4%Aut.
MATH83.8%Aut.
MathVista-Mini62.9%Aut.
HiddenMath54.5%Aut.

Multimodal

AI2D84.2%Aut.
ChartQAMasry et al. (2022)75.7%Aut.
InfoVQA64.9%Aut.

Reasoning

Bird-SQL (dev)47.9%Aut.

Índices de evaluación AA

(Artificial Analysis)
Math Index(Artificial Analysis)
18.3
Coding Index(Artificial Analysis)
5.8
Intelligence Index(Artificial Analysis)
5.5
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.6
Ifbench(Google Research (2023))
0.4
Gpqa(NYU + Cohere + Anthropic (2023))
0.3
Aime(MAA (Mathematical Association of America))
0.2
Aime 25(MAA (Mathematical Association of America))
0.2
Scicode(UIUC + Argonne National Lab (2024))
0.2
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.1
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.1
Lcr(Artificial Analysis)
0.1
Hle(Center for AI Safety + Scale AI (2025))
0.0
Tau Banking
0.0
Terminalbench Hard(Stanford × Laude Institute (2026))
0.0
Terminalbench V2 1
0.0

Puntuaciones por categoría LLM Stats

(LLM Stats (zeroeval))
Structured Output
90
Instruction Following
90
Image To Text
80
Grounding
80
Math
70
Multimodal
70
Vision
70
Legal
60
Reasoning
60
Finance
60
General
60
Healthcare
60
Code
60
Language
50
Physics
40
Factuality
40
Biology
40
Chemistry
40

Precios

Precio de entradaGratis
Precio de salidaGratis
Precio mixto (3:1)Gratis

Velocidad

Tokens/seg0.0
Retraso del primer token0.00s
Tiempo hasta la respuesta0.00s

Ranking de Precios por Proveedor

Ranking de Precios por Proveedor

1 proveedores

ProveedorEntradaSalida
1Neon
$0.15
$0.5

Comparar precios entre diferentes proveedores de API para este modelo.

Fuentes externas