Gemma 3 12B Instruct
GoogleGemmaOpen WeightGemma · Commercial OK
Description
Gemma 3 12B is a 12-billion-parameter vision-language model from Google, handling text and image input and generating text output. It features a 128K context window, multilingual support, and open weights. Suitable for question answering, summarization, reasoning, and image understanding tasks.
Release Date
2025-03-12
Parameters
12.0B
Context Length
131K
Modalities
image, text
Capability Radar
22
general
10
coding
31
reasoning
23
science
25
agents
80
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 519 | 8.0 | AA |
| General Ranking | 474 | 25.0 | AA |
| Multimodal Ranking | 57 | 43.0 | LS |
| Science | 501 | 21.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
40.9%SR
Code
HumanEvalOpenAI (2021)
85.4%SR
LiveCodeBench
24.6%SR
Factuality
FACTS Grounding
75.8%SR
SimpleQA
6.3%SR
Finance
MMLU-Pro
60.6%SR
General
IFEvalGoogle Research (2023)
88.9%SR
Natural2Code
80.7%SR
MBPP
0.73 / 100SR
Global-MMLU-Lite
69.5%SR
MMMU (val)
59.6%SR
BIG-Bench Extra Hard
16.3%SR
Image To Text
DocVQADocVQA (2020)
87.1%SR
VQAv2 (val)
71.6%SR
TextVQA
67.7%SR
Language
BIG-Bench Hard
85.7%SR
WMT24++
51.6%SR
ECLeKTic
10.3%SR
Math
GSM8k
94.4%SR
MATH
83.8%SR
MathVista-Mini
62.9%SR
HiddenMath
54.5%SR
Multimodal
AI2D
84.2%SR
ChartQAMasry et al. (2022)
75.7%SR
InfoVQA
64.9%SR
Reasoning
Bird-SQL (dev)
47.9%SR
AA Evaluation Indices
(Artificial Analysis)Math Index(Artificial Analysis)18.3
Coding Index(Artificial Analysis)5.8
Intelligence Index(Artificial Analysis)5.5
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.9
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.6
Ifbench(Google Research (2023))0.4
Gpqa(NYU + Cohere + Anthropic (2023))0.3
Aime(MAA (Mathematical Association of America))0.2
Aime 25(MAA (Mathematical Association of America))0.2
Scicode(UIUC + Argonne National Lab (2024))0.2
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.1
Tau2(Sierra + U Toronto + Vector Institute (2025))0.1
Lcr(Artificial Analysis)0.1
Hle(Center for AI Safety + Scale AI (2025))0.0
Tau Banking0.0
Terminalbench Hard(Stanford × Laude Institute (2026))0.0
Terminalbench V2 10.0
LLM Stats Category Scores
(LLM Stats (zeroeval))Structured Output90
Instruction Following90
Image To Text80
Grounding80
Math70
Multimodal70
Vision70
Legal60
Reasoning60
Finance60
General60
Healthcare60
Code60
Language50
Physics40
Factuality40
Biology40
Chemistry40
Pricing
Input PriceFree
Output PriceFree
Blended Price (3:1)Free
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
Provider Price Ranking
1 providers
ProviderInputOutput
1Neon
$0.15
$0.5
Compare pricing across different API providers for this model.