Skip to main content

Gemma 3 27B Instruct

GoogleGemmaOpen WeightGemma · Commercial OK

Description

Gemma 3 27B is a 27-billion-parameter vision-language model from Google, handling text and image input and generating text output. It features a 128K context window, multilingual support, and open weights. Suitable for complex question answering, summarization, reasoning, and image understanding tasks.

Release Date
2025-03-12
Parameters
27.0B
Context Length
131K
Modalities
image, text

Capability Radar

23
general
13
coding
34
reasoning
28
science
28
agents
70
multimodal

Rankings

Domain#RankScoreSource
Code Ranking578
11.0
AA
General Ranking549
24.0
AA
Multimodal Ranking109
41.0
LS
Science550
23.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Chat

IFEvalGoogle Research (2023)90.4%SR

Factuality

SimpleQA10.0%SR

General

Global-MMLU-Lite75.1%SR

Language

MMLU-Pro67.5%SR
WMT24++53.4%SR
ECLeKTic16.7%SR

Long Context

MRCR v2 (8-needle)13.5%SR

Math

GSM8k95.9%SR
MATH89.0%SR
MathVista-Mini67.6%SR
HiddenMath60.3%SR

Reasoning

HumanEvalOpenAI (2021)87.8%SR
BIG-Bench Hard87.6%SR
Natural2Code84.5%SR
ChartQAMasry et al. (2022)78.0%SR
FACTS Grounding74.9%SR
MBPP0.74 / 100SR
Bird-SQL (dev)54.4%SR
GPQANYU + Cohere + Anthropic (2023)42.4%SR
LiveCodeBench29.7%SR
BIG-Bench Extra Hard19.3%SR

Vision

DocVQADocVQA (2020)86.6%SR
AI2D84.5%SR
VQAv2 (val)71.0%SR
InfoVQA70.6%SR
TextVQA65.1%SR
MMMU (val)64.9%SR

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
88.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
66.9
Gpqa(NYU + Cohere + Anthropic (2023))
42.8
Ifbench(Google Research (2023))
31.8
Aime(MAA (Mathematical Association of America))
25.3
Scicode(UIUC + Argonne National Lab (2024))
23.3
Math Index(Artificial Analysis)
20.7
Aime 25(MAA (Mathematical Association of America))
20.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
13.7
Tau2(Sierra + U Toronto + Vector Institute (2025))
10.5
Coding Index(Artificial Analysis)
10.1
Lcr(Artificial Analysis)
7.3
Intelligence Index(Artificial Analysis)
4.9
Terminalbench V2 1
4.5
Hle(Center for AI Safety + Scale AI (2025))
4.4
Terminalbench Hard(Stanford × Laude Institute (2026))
3.8
Tau Banking
0.8
Terminalbench V4 0
0.0

LLM Stats Category Scores

(LLM Stats (zeroeval))
Chat
90
Instruction Following
90
Structured Output
90
Math
80
Image To Text
70
Legal
70
Multimodal
70
Finance
70
Grounding
70
Healthcare
70
Vision
70
Language
60
Reasoning
60
General
60
Code
60
Physics
40
Factuality
40
Biology
40
Chemistry
40
Long Context
10

Pricing

Input PriceFree
Output PriceFree
Blended Price (3:1)Free
Cache Read Price$0.04 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

9 providers

Cheapest: DeepInfraMost Expensive: STACKIT
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2OpenRouter
$0.08
$0.45
3Hugging Face
$0.08
$0.16
4Deep Infra
$0.08
$0.16
5Kilo Gateway
$0.08
$0.16
6Merge Gateway
$0.08
$0.45
7Nebius Token Factory
$0.1
$0.3
8NovitaAI
$0.119
$0.2
9STACKIT
$0.53
$0.76

Compare pricing across different API providers for this model.

External Sources