メインコンテンツへスキップ

Phi-4 Mini Instruct

MicrosoftPhiオープンウエイトMIT · 商用利用可

説明

Phi 4 Mini Instruct is a lightweight (3.8B parameters) open model built upon synthetic data and filtered web data, focusing on high-quality reasoning. It supports a 128K token context length and is enhanced for instruction adherence and safety via supervised fine-tuning and direct preference optimization.

リリース日
2024-02-26
パラメータ
3.8B
コンテキスト長
モダリティ

能力レーダー

18
general
8
coding
18
reasoning
20
science
15
agents
0
multimodal

ランキング

ドメイン#順位スコアソース
コーディングランキング511
9.0
AA
総合ランキング522
17.0
AA
科学513
17.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Biology

GPQANYU + Cohere + Anthropic (2023)52.0%自己申告

Creativity

Social IQa72.5%自己申告
Arena Hard32.8%自己申告

Finance

MMLU67.3%自己申告
TruthfulQA66.4%自己申告
MMLU-Pro52.8%自己申告

General

ARC-C83.7%自己申告
OpenBookQA79.2%自己申告
PIQA77.6%自己申告
Multilingual MMLU49.3%自己申告

Language

BoolQ81.2%自己申告
BIG-Bench Hard70.4%自己申告
Winogrande67.0%自己申告

Math

MATH-50094.6%自己申告
GSM8k88.6%自己申告
MATH64.0%自己申告
MGSM63.9%自己申告
AIMEMAA57.5%自己申告

Reasoning

HellaSwagAI2 (2019)69.1%自己申告

AA評価指数

(Artificial Analysis)
Math Index(Artificial Analysis)
6.7
Intelligence Index(Artificial Analysis)
5.7
Coding Index(Artificial Analysis)
3.8
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
0.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.5
Gpqa(NYU + Cohere + Anthropic (2023))
0.3
Ifbench(Google Research (2023))
0.2
Lcr(Artificial Analysis)
0.2
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.1
Scicode(UIUC + Argonne National Lab (2024))
0.1
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.1
Aime 25(MAA (Mathematical Association of America))
0.1
Hle(Center for AI Safety + Scale AI (2025))
0.0
Aime(MAA (Mathematical Association of America))
0.0
Terminalbench V2 1
0.0
Terminalbench Hard(Stanford × Laude Institute (2026))
0.0

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Math
70
Psychology
70
Reasoning
70
General
70
Language
60
Legal
60
Finance
60
Healthcare
60
Physics
50
Creativity
50
Biology
30
Chemistry
30
Writing
30

価格設定

入力価格無料
出力価格無料
混合価格(3:1)無料

速度

トークン/秒45.2
初トークン遅延0.36s
初回答遅延0.36s

プロバイダー価格ランキング

プロバイダー価格ランキング

2 プロバイダー

最安: Azure Cognitive Services最高: Azure
プロバイダー入力出力
1Azure Cognitive Services最安
$0.075
$0.3
2Azure
$0.075
$0.3

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク