Gemma 3 4B Instruct
GoogleGemmaオープンウエイトGemma · 商用利用可
説明
Gemma 3 4B is a 4-billion-parameter vision-language model from Google, handling text and image input and generating text output. It features a 128K context window, multilingual support, and open weights. Suitable for question answering, summarization, reasoning, and image understanding tasks.
リリース日
2025-03-12
パラメータ
4.0B
コンテキスト長
131K
モダリティ
image, text
能力レーダー
16
general
6
coding
22
reasoning
22
science
17
agents
70
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| コーディングランキング | 618 | 5.0 | AA |
| 総合ランキング | 633 | 15.0 | AA |
| マルチモーダルランキング | 133 | 32.0 | LS |
| 科学 | 625 | 15.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Chat
IFEvalGoogle Research (2023)
90.2%自己申告
Factuality
SimpleQA
4.0%自己申告
General
Global-MMLU-Lite
54.5%自己申告
Language
WMT24++
46.8%自己申告
MMLU-Pro
43.6%自己申告
ECLeKTic
4.6%自己申告
Math
GSM8k
89.2%自己申告
MATH
75.6%自己申告
MathVista-Mini
50.0%自己申告
HiddenMath
43.0%自己申告
Reasoning
BIG-Bench Hard
72.2%自己申告
HumanEvalOpenAI (2021)
71.3%自己申告
Natural2Code
70.3%自己申告
FACTS Grounding
70.1%自己申告
ChartQAMasry et al. (2022)
68.8%自己申告
MBPP
0.63 / 100自己申告
Bird-SQL (dev)
36.3%自己申告
GPQANYU + Cohere + Anthropic (2023)
30.8%自己申告
LiveCodeBench
12.6%自己申告
BIG-Bench Extra Hard
11.0%自己申告
Vision
DocVQADocVQA (2020)
75.8%自己申告
AI2D
74.8%自己申告
VQAv2 (val)
62.4%自己申告
TextVQA
57.8%自己申告
InfoVQA
50.0%自己申告
MMMU (val)
48.8%自己申告
AA評価指数
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))76.6
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))41.7
Gpqa(NYU + Cohere + Anthropic (2023))29.1
Ifbench(Google Research (2023))28.3
Math Index(Artificial Analysis)12.7
Aime 25(MAA (Mathematical Association of America))12.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))11.2
Lcr(Artificial Analysis)6.7
Aime(MAA (Mathematical Association of America))6.3
Hle(Center for AI Safety + Scale AI (2025))5.3
Tau2(Sierra + U Toronto + Vector Institute (2025))5.0
Intelligence Index(Artificial Analysis)4.8
Coding Index(Artificial Analysis)2.7
Terminalbench Hard(Stanford × Laude Institute (2026))0.8
Tau Banking0.4
Terminalbench V2 10.4
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Chat90
Instruction Following90
Structured Output90
Image To Text70
Grounding70
Math60
Multimodal60
Vision60
Reasoning50
General50
Healthcare50
Language40
Legal40
Factuality40
Finance40
Code40
Physics30
Biology30
Chemistry30
価格設定
入力価格無料
出力価格無料
混合価格(3:1)無料
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
6 プロバイダー
最安: DeepInfra最高: Kilo Gateway
プロバイダー入力出力
1DeepInfra最安
$0
$0
2Merge Gateway
$0.04
$0.08
3OpenRouter
$0.05
$0.1
4Hugging Face
$0.05
$0.1
5Deep Infra
$0.05
$0.1
6Kilo Gateway
$0.05
$0.1
このモデルの異なるAPIプロバイダー間の価格を比較。