メインコンテンツへスキップ

Granite 3.3 8B (Non-reasoning)

IBMオープンウエイトApache 2.0 · 商用利用可

説明

Granite-3.3-8B-Base is a decoder-only language model with a 128K token context window. It improves upon Granite-3.1-8B-Base by adding support for Fill-in-the-Middle (FIM) using specialized tokens, enabling the model to generate content conditioned on both prefix and suffix. This makes it well-suited for code completion tasks

リリース日
2025-04-16
パラメータ
8.2B
コンテキスト長
—
モダリティ
text

能力レーダー

17
general
13
coding
18
reasoning
25
science
17
agents
0
multimodal

ランキング

ドメイン#順位スコアソース
コーディングランキング609
7.0
AA
総合ランキング620
16.0
AA
科学596
17.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Chat

IFEvalGoogle Research (2023)74.8%自己申告

General

TriviaQA78.2%自己申告
TruthfulQA66.9%自己申告
MMLU65.5%自己申告
AlpacaEval 2.062.7%自己申告
Arena Hard57.6%自己申告
PopQA26.2%自己申告

Math

AIME 202481.2%自己申告
GSM8k80.9%自己申告
MATH-50069.0%自己申告

Reasoning

HumanEvalOpenAI (2021)89.7%自己申告
HumanEval+86.1%自己申告
HellaSwagAI2 (2019)80.1%自己申告
Winogrande74.4%自己申告
BIG-Bench Hard69.1%自己申告
DROP59.4%自己申告
ARC-C50.8%自己申告
AGIEval49.3%自己申告
NQ36.5%自己申告

Safety

AttaQ88.5%自己申告

AA評価指数

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
66.5
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
46.8
Gpqa(NYU + Cohere + Anthropic (2023))
33.8
Ifbench(Google Research (2023))
22.4
Livecodebench(UC Berkeley + MIT + Cornell (2024))
12.7
Tau2(Sierra + U Toronto + Vector Institute (2025))
10.5
Math Index(Artificial Analysis)
6.7
Aime 25(MAA (Mathematical Association of America))
6.7
Intelligence Index(Artificial Analysis)
4.9
Aime(MAA (Mathematical Association of America))
4.7
Hle(Center for AI Safety + Scale AI (2025))
4.2
Terminalbench Hard(Stanford × Laude Institute (2026))
0.0

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Safety
90
Code
90
Chat
70
Instruction Following
70
Language
70
Structured Output
70
General
70
Legal
60
Math
60
Reasoning
60
Finance
60
Healthcare
60
Creativity
60
Writing
60

価格設定

入力価格$0.03 / 1Mトークン
出力価格$0.25 / 1Mトークン
混合価格(3:1)$0.085 / 1Mトークン

速度

トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s

プロバイダー価格ランキング

プロバイダー価格ランキング

1 プロバイダー

プロバイダー入力出力
1IBMプライマリ
$0.03
$0.25

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク