메인 콘텐츠로 건너뛰기

Phi-4 Mini Instruct

MicrosoftPhi오픈 웨이트MIT · 상업적 사용 가능

설명

Phi 4 Mini Instruct is a lightweight (3.8B parameters) open model built upon synthetic data and filtered web data, focusing on high-quality reasoning. It supports a 128K token context length and is enhanced for instruction adherence and safety via supervised fine-tuning and direct preference optimization.

출시일
2024-02-26
파라미터
3.8B
컨텍스트 길이
모달리티

능력 레이더

18
general
8
coding
18
reasoning
20
science
15
agents
0
multimodal

랭킹

도메인#순위점수소스
코딩 랭킹504
9.0
AA
종합 랭킹515
17.0
AA
과학506
17.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Biology

GPQANYU + Cohere + Anthropic (2023)52.0%자체 보고

Creativity

Social IQa72.5%자체 보고
Arena Hard32.8%자체 보고

Finance

MMLU67.3%자체 보고
TruthfulQA66.4%자체 보고
MMLU-Pro52.8%자체 보고

General

ARC-C83.7%자체 보고
OpenBookQA79.2%자체 보고
PIQA77.6%자체 보고
Multilingual MMLU49.3%자체 보고

Language

BoolQ81.2%자체 보고
BIG-Bench Hard70.4%자체 보고
Winogrande67.0%자체 보고

Math

MATH-50094.6%자체 보고
GSM8k88.6%자체 보고
MATH64.0%자체 보고
MGSM63.9%자체 보고
AIMEMAA57.5%자체 보고

Reasoning

HellaSwagAI2 (2019)69.1%자체 보고

AA 평가 지수

(Artificial Analysis)
Math Index(Artificial Analysis)
6.7
Intelligence Index(Artificial Analysis)
5.7
Coding Index(Artificial Analysis)
3.8
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
0.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.5
Gpqa(NYU + Cohere + Anthropic (2023))
0.3
Ifbench(Google Research (2023))
0.2
Lcr(Artificial Analysis)
0.2
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.1
Scicode(UIUC + Argonne National Lab (2024))
0.1
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.1
Aime 25(MAA (Mathematical Association of America))
0.1
Hle(Center for AI Safety + Scale AI (2025))
0.0
Aime(MAA (Mathematical Association of America))
0.0
Terminalbench V2 1
0.0
Terminalbench Hard(Stanford × Laude Institute (2026))
0.0

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Math
70
Psychology
70
Reasoning
70
General
70
Language
60
Legal
60
Finance
60
Healthcare
60
Physics
50
Creativity
50
Biology
30
Chemistry
30
Writing
30

가격

입력 가격무료
출력 가격무료
혼합 가격 (3:1)무료

속도

토큰/초45.0
첫 토큰 지연0.34s
첫 응답 지연0.34s

공급자 가격 순위

공급자 가격 순위

2개 공급자

최저가: Azure Cognitive Services최고가: Azure
공급자입력출력
1Azure Cognitive Services최저가
$0.075
$0.3
2Azure
$0.075
$0.3

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크