Gemma 3 4B Instruct
GoogleGemmaOpen WeightGemma · Usage Commercial
Description
Gemma 3 4B is a 4-billion-parameter vision-language model from Google, handling text and image input and generating text output. It features a 128K context window, multilingual support, and open weights. Suitable for question answering, summarization, reasoning, and image understanding tasks.
Date de sortie
2025-03-12
Paramètres
4.0B
Longueur du contexte
131K
Modalités
image, text
Radar de capacités
14
general
6
coding
22
reasoning
17
science
17
agents
70
multimodal
Classements
| Domaine | #Rang | Score | Source |
|---|---|---|---|
| Classement codage | 533 | 5.0 | AA |
| Classement général | 550 | 14.0 | AA |
| Classement multimodal | 70 | 36.0 | LS |
| Science | 543 | 14.0 | AA |
Scores de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
30.8%Aut.
Code
HumanEvalOpenAI (2021)
71.3%Aut.
LiveCodeBench
12.6%Aut.
Factuality
FACTS Grounding
70.1%Aut.
SimpleQA
4.0%Aut.
Finance
MMLU-Pro
43.6%Aut.
General
IFEvalGoogle Research (2023)
90.2%Aut.
Natural2Code
70.3%Aut.
MBPP
0.63 / 100Aut.
Global-MMLU-Lite
54.5%Aut.
MMMU (val)
48.8%Aut.
BIG-Bench Extra Hard
11.0%Aut.
Image To Text
DocVQADocVQA (2020)
75.8%Aut.
VQAv2 (val)
62.4%Aut.
TextVQA
57.8%Aut.
Language
BIG-Bench Hard
72.2%Aut.
WMT24++
46.8%Aut.
ECLeKTic
4.6%Aut.
Math
GSM8k
89.2%Aut.
MATH
75.6%Aut.
MathVista-Mini
50.0%Aut.
HiddenMath
43.0%Aut.
Multimodal
AI2D
74.8%Aut.
ChartQAMasry et al. (2022)
68.8%Aut.
InfoVQA
50.0%Aut.
Reasoning
Bird-SQL (dev)
36.3%Aut.
Indices d'évaluation AA
(Artificial Analysis)Math Index(Artificial Analysis)12.7
Coding Index(Artificial Analysis)2.7
Intelligence Index(Artificial Analysis)1.0
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.8
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.4
Gpqa(NYU + Cohere + Anthropic (2023))0.3
Ifbench(Google Research (2023))0.3
Aime 25(MAA (Mathematical Association of America))0.1
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.1
Scicode(UIUC + Argonne National Lab (2024))0.1
Lcr(Artificial Analysis)0.1
Aime(MAA (Mathematical Association of America))0.1
Hle(Center for AI Safety + Scale AI (2025))0.1
Tau2(Sierra + U Toronto + Vector Institute (2025))0.0
Terminalbench Hard(Stanford × Laude Institute (2026))0.0
Tau Banking0.0
Terminalbench V2 10.0
Scores par catégorie LLM Stats
(LLM Stats (zeroeval))Structured Output90
Instruction Following90
Image To Text70
Grounding70
Math60
Multimodal60
Vision60
Reasoning50
General50
Healthcare50
Legal40
Language40
Factuality40
Finance40
Code40
Physics30
Biology30
Chemistry30
Tarification
Prix d'entréeGratuit
Prix de sortieGratuit
Prix mixte (3:1)Gratuit
Vitesse
Tokens/sec0.0
Délai du premier token0.00s
Temps de réponse0.00s
Classement des Prix par Fournisseur
Classement des Prix par Fournisseur
2 fournisseurs
Moins cher: OpenRouterPlus cher: Kilo Gateway
FournisseurEntréeSortie
1OpenRouterMoins cher
$0.05
$0.1
2Kilo Gateway
$0.05
$0.1
Comparer les prix entre différents fournisseurs API pour ce modèle.