Gemma 3 27B Instruct
GoogleGemmaOpen WeightGemma · Commercial OK
Description
Gemma 3 27B is a 27-billion-parameter vision-language model from Google, handling text and image input and generating text output. It features a 128K context window, multilingual support, and open weights. Suitable for complex question answering, summarization, reasoning, and image understanding tasks.
Release Date
2025-03-12
Parameters
27.0B
Context Length
131K
Modalities
image, text
Capability Radar
23
general
13
coding
34
reasoning
28
science
28
agents
70
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 578 | 11.0 | AA |
| General Ranking | 549 | 24.0 | AA |
| Multimodal Ranking | 109 | 41.0 | LS |
| Science | 550 | 23.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Chat
IFEvalGoogle Research (2023)
90.4%SR
Factuality
SimpleQA
10.0%SR
General
Global-MMLU-Lite
75.1%SR
Language
MMLU-Pro
67.5%SR
WMT24++
53.4%SR
ECLeKTic
16.7%SR
Long Context
MRCR v2 (8-needle)
13.5%SR
Math
GSM8k
95.9%SR
MATH
89.0%SR
MathVista-Mini
67.6%SR
HiddenMath
60.3%SR
Reasoning
HumanEvalOpenAI (2021)
87.8%SR
BIG-Bench Hard
87.6%SR
Natural2Code
84.5%SR
ChartQAMasry et al. (2022)
78.0%SR
FACTS Grounding
74.9%SR
MBPP
0.74 / 100SR
Bird-SQL (dev)
54.4%SR
GPQANYU + Cohere + Anthropic (2023)
42.4%SR
LiveCodeBench
29.7%SR
BIG-Bench Extra Hard
19.3%SR
Vision
DocVQADocVQA (2020)
86.6%SR
AI2D
84.5%SR
VQAv2 (val)
71.0%SR
InfoVQA
70.6%SR
TextVQA
65.1%SR
MMMU (val)
64.9%SR
AA Evaluation Indices
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))88.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))66.9
Gpqa(NYU + Cohere + Anthropic (2023))42.8
Ifbench(Google Research (2023))31.8
Aime(MAA (Mathematical Association of America))25.3
Scicode(UIUC + Argonne National Lab (2024))23.3
Math Index(Artificial Analysis)20.7
Aime 25(MAA (Mathematical Association of America))20.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))13.7
Tau2(Sierra + U Toronto + Vector Institute (2025))10.5
Coding Index(Artificial Analysis)10.1
Lcr(Artificial Analysis)7.3
Intelligence Index(Artificial Analysis)4.9
Terminalbench V2 14.5
Hle(Center for AI Safety + Scale AI (2025))4.4
Terminalbench Hard(Stanford × Laude Institute (2026))3.8
Tau Banking0.8
Terminalbench V4 00.0
LLM Stats Category Scores
(LLM Stats (zeroeval))Chat90
Instruction Following90
Structured Output90
Math80
Image To Text70
Legal70
Multimodal70
Finance70
Grounding70
Healthcare70
Vision70
Language60
Reasoning60
General60
Code60
Physics40
Factuality40
Biology40
Chemistry40
Long Context10
Pricing
Input PriceFree
Output PriceFree
Blended Price (3:1)Free
Cache Read Price$0.04 / 1M tokens
Speed
Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s
Provider Price Ranking
Provider Price Ranking
9 providers
Cheapest: DeepInfraMost Expensive: STACKIT
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2OpenRouter
$0.08
$0.45
3Hugging Face
$0.08
$0.16
4Deep Infra
$0.08
$0.16
5Kilo Gateway
$0.08
$0.16
6Merge Gateway
$0.08
$0.45
7Nebius Token Factory
$0.1
$0.3
8NovitaAI
$0.119
$0.2
9STACKIT
$0.53
$0.76
Compare pricing across different API providers for this model.