メインコンテンツへスキップ

Gemma 3 4B Instruct

GoogleGemmaオープンウエイトGemma · 商用利用可

説明

Gemma 3 4B is a 4-billion-parameter vision-language model from Google, handling text and image input and generating text output. It features a 128K context window, multilingual support, and open weights. Suitable for question answering, summarization, reasoning, and image understanding tasks.

リリース日
2025-03-12
パラメータ
4.0B
コンテキスト長
131K
モダリティ
image, text

能力レーダー

16
general
6
coding
22
reasoning
22
science
17
agents
70
multimodal

ランキング

ドメイン#順位スコアソース
コーディングランキング618
5.0
AA
総合ランキング633
15.0
AA
マルチモーダルランキング133
32.0
LS
科学625
15.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Chat

IFEvalGoogle Research (2023)90.2%自己申告

Factuality

SimpleQA4.0%自己申告

General

Global-MMLU-Lite54.5%自己申告

Language

WMT24++46.8%自己申告
MMLU-Pro43.6%自己申告
ECLeKTic4.6%自己申告

Math

GSM8k89.2%自己申告
MATH75.6%自己申告
MathVista-Mini50.0%自己申告
HiddenMath43.0%自己申告

Reasoning

BIG-Bench Hard72.2%自己申告
HumanEvalOpenAI (2021)71.3%自己申告
Natural2Code70.3%自己申告
FACTS Grounding70.1%自己申告
ChartQAMasry et al. (2022)68.8%自己申告
MBPP0.63 / 100自己申告
Bird-SQL (dev)36.3%自己申告
GPQANYU + Cohere + Anthropic (2023)30.8%自己申告
LiveCodeBench12.6%自己申告
BIG-Bench Extra Hard11.0%自己申告

Vision

DocVQADocVQA (2020)75.8%自己申告
AI2D74.8%自己申告
VQAv2 (val)62.4%自己申告
TextVQA57.8%自己申告
InfoVQA50.0%自己申告
MMMU (val)48.8%自己申告

AA評価指数

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
76.6
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
41.7
Gpqa(NYU + Cohere + Anthropic (2023))
29.1
Ifbench(Google Research (2023))
28.3
Math Index(Artificial Analysis)
12.7
Aime 25(MAA (Mathematical Association of America))
12.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))
11.2
Lcr(Artificial Analysis)
6.7
Aime(MAA (Mathematical Association of America))
6.3
Hle(Center for AI Safety + Scale AI (2025))
5.3
Tau2(Sierra + U Toronto + Vector Institute (2025))
5.0
Intelligence Index(Artificial Analysis)
4.8
Coding Index(Artificial Analysis)
2.7
Terminalbench Hard(Stanford × Laude Institute (2026))
0.8
Tau Banking
0.4
Terminalbench V2 1
0.4

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Chat
90
Instruction Following
90
Structured Output
90
Image To Text
70
Grounding
70
Math
60
Multimodal
60
Vision
60
Reasoning
50
General
50
Healthcare
50
Language
40
Legal
40
Factuality
40
Finance
40
Code
40
Physics
30
Biology
30
Chemistry
30

価格設定

入力価格無料
出力価格無料
混合価格(3:1)無料

速度

トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s

プロバイダー価格ランキング

プロバイダー価格ランキング

6 プロバイダー

最安: DeepInfra最高: Kilo Gateway
プロバイダー入力出力
1DeepInfra最安
$0
$0
2Merge Gateway
$0.04
$0.08
3OpenRouter
$0.05
$0.1
4Hugging Face
$0.05
$0.1
5Deep Infra
$0.05
$0.1
6Kilo Gateway
$0.05
$0.1

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク