Granite 4.0 Micro
IBMOpen WeightApache 2.0 · Usage Commercial
Description
A preliminary version of the smallest model in the upcoming Granite 4.0 family, released May 2025. It utilizes a novel hybrid Mamba-2/Transformer, fine-grained mixture of experts (MoE) architecture (7B total parameters, 1B active at inference). This preview version is partially trained (2.5T tokens) but demonstrates significant memory efficiency and performance potential, validated for at least 128K context length without positional encoding.
Date de sortie
2025-09-22
Paramètres
7.0B
Longueur du contexte
131K
Modalités
text
Radar de capacités
15
general
17
coding
11
reasoning
20
science
13
agents
0
multimodal
Classements
| Domaine | #Rang | Score | Source |
|---|---|---|---|
| Classement codage | 506 | 9.0 | AA |
| Classement général | 527 | 16.0 | AA |
| Science | 502 | 18.0 | AA |
Scores de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Code
HumanEvalOpenAI (2021)
82.4%Aut.
Creativity
AlpacaEval 2.0
35.2%Aut.
Arena Hard
26.7%Aut.
Finance
MMLU
60.4%Aut.
TruthfulQA
58.1%Aut.
General
IFEvalGoogle Research (2023)
63.0%Aut.
PopQA
22.9%Aut.
Language
BIG-Bench Hard
55.7%Aut.
Math
GSM8k
70.1%Aut.
DROP
46.2%Aut.
Reasoning
HumanEval+
78.3%Aut.
Safety
AttaQ
86.1%Aut.
Indices d'évaluation AA
(Artificial Analysis)Math Index(Artificial Analysis)6.0
Intelligence Index(Artificial Analysis)2.0
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.4
Gpqa(NYU + Cohere + Anthropic (2023))0.3
Ifbench(Google Research (2023))0.2
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.2
Tau2(Sierra + U Toronto + Vector Institute (2025))0.1
Scicode(UIUC + Argonne National Lab (2024))0.1
Aime 25(MAA (Mathematical Association of America))0.1
Lcr(Artificial Analysis)0.1
Hle(Center for AI Safety + Scale AI (2025))0.1
Terminalbench Hard(Stanford × Laude Institute (2026))0.0
Scores par catégorie LLM Stats
(LLM Stats (zeroeval))Safety90
Code80
Legal60
Math60
Structured Output60
Instruction Following60
Language60
Finance60
General60
Healthcare60
Reasoning50
Creativity30
Writing30
Tarification
Prix d'entréeGratuit
Prix de sortieGratuit
Prix mixte (3:1)Gratuit
Vitesse
Tokens/sec0.0
Délai du premier token0.00s
Temps de réponse0.00s
Classement des Prix par Fournisseur
Classement des Prix par Fournisseur
2 fournisseurs
Moins cher: OpenRouterPlus cher: Kilo Gateway
FournisseurEntréeSortie
1OpenRouterMoins cher
$0.017
$0.112
2Kilo Gateway
$0.017
$0.112
Comparer les prix entre différents fournisseurs API pour ce modèle.