मुख्य सामग्री पर जाएं

Phi-4 Mini Instruct

MicrosoftPhiओपन वेटMIT · व्यावसायिक उपयोग

विवरण

Phi 4 Mini Instruct is a lightweight (3.8B parameters) open model built upon synthetic data and filtered web data, focusing on high-quality reasoning. It supports a 128K token context length and is enhanced for instruction adherence and safety via supervised fine-tuning and direct preference optimization.

रिलीज़ तिथि
2024-02-26
पैरामीटर
3.8B
संदर्भ लंबाई
—
मोडैलिटीज़
—

क्षमता रडार

18
general
7
coding
18
reasoning
24
science
15
agents
0
multimodal

रैंकिंग

डोमेन#रैंकस्कोरस्रोत
कोडिंग रैंकिंग592
8.0
AA
सामान्य रैंकिंग611
16.0
AA
विज्ञान594
16.0
AA

बेंचमार्क स्कोर (LLM Stats)

(LLM Stats (zeroeval))

General

MMLU67.3%स्वयं
TruthfulQA66.4%स्वयं
Multilingual MMLU49.3%स्वयं
Arena Hard32.8%स्वयं

Language

BoolQ81.2%स्वयं
MMLU-Pro52.8%स्वयं

Math

MATH-50094.6%स्वयं
GSM8k88.6%स्वयं
MATH64.0%स्वयं
MGSM63.9%स्वयं
AIMEMAA57.5%स्वयं

Reasoning

ARC-C83.7%स्वयं
OpenBookQA79.2%स्वयं
PIQA77.6%स्वयं
Social IQa72.5%स्वयं
BIG-Bench Hard70.4%स्वयं
HellaSwagAI2 (2019)69.1%स्वयं
Winogrande67.0%स्वयं
GPQANYU + Cohere + Anthropic (2023)52.0%स्वयं

AA मूल्यांकन सूचकांक

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
69.6
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
46.5
Gpqa(NYU + Cohere + Anthropic (2023))
33.1
Ifbench(Google Research (2023))
21.1
Lcr(Artificial Analysis)
15.3
Livecodebench(UC Berkeley + MIT + Cornell (2024))
12.6
Tau2(Sierra + U Toronto + Vector Institute (2025))
8.2
Math Index(Artificial Analysis)
6.7
Aime 25(MAA (Mathematical Association of America))
6.7
Intelligence Index(Artificial Analysis)
6.3
Hle(Center for AI Safety + Scale AI (2025))
4.4
Coding Index(Artificial Analysis)
3.8
Aime(MAA (Mathematical Association of America))
3.0
Terminalbench V2 1
0.4
Terminalbench Hard(Stanford × Laude Institute (2026))
0.0

LLM Stats श्रेणी स्कोर

(LLM Stats (zeroeval))
Math
70
Psychology
70
Reasoning
70
General
70
Language
60
Legal
60
Finance
60
Healthcare
60
Physics
50
Creativity
50
Chat
30
Biology
30
Chemistry
30
Writing
30

मूल्य निर्धारण

इनपुट मूल्यमुफ्त
आउटपुट मूल्यमुफ्त
मिश्रित मूल्य (3:1)मुफ्त

गति

टोकन/सेकंड47.6
पहले टोकन में देरी0.33s
पहले उत्तर में देरी0.33s

प्रदाता मूल्य रैंकिंग

प्रदाता मूल्य रैंकिंग

2 प्रदाता

सबसे सस्ता: Azure Cognitive Servicesसबसे महंगा: Azure
प्रदाताइनपुटआउटपुट
1Azure Cognitive Servicesसबसे सस्ता
$0.075
$0.3
2Azure
$0.075
$0.3

इस मॉडल के लिए विभिन्न API प्रदाताओं के मूल्य निर्धारण की तुलना करें।

बाहरी लिंक