Phi-4 Mini Instruct
MicrosoftPhiओपन वेटMIT · व्यावसायिक उपयोग
विवरण
Phi 4 Mini Instruct is a lightweight (3.8B parameters) open model built upon synthetic data and filtered web data, focusing on high-quality reasoning. It supports a 128K token context length and is enhanced for instruction adherence and safety via supervised fine-tuning and direct preference optimization.
रिलीज़ तिथि
2024-02-26
पैरामीटर
3.8B
संदर्भ लंबाई
—
मोडैलिटीज़
—
क्षमता रडार
18
general
8
coding
18
reasoning
20
science
15
agents
0
multimodal
रैंकिंग
| डोमेन | #रैंक | स्कोर | स्रोत |
|---|---|---|---|
| कोडिंग रैंकिंग | 504 | 9.0 | AA |
| सामान्य रैंकिंग | 515 | 17.0 | AA |
| विज्ञान | 507 | 17.0 | AA |
बेंचमार्क स्कोर (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
52.0%स्वयं
Creativity
Social IQa
72.5%स्वयं
Arena Hard
32.8%स्वयं
Finance
MMLU
67.3%स्वयं
TruthfulQA
66.4%स्वयं
MMLU-Pro
52.8%स्वयं
General
ARC-C
83.7%स्वयं
OpenBookQA
79.2%स्वयं
PIQA
77.6%स्वयं
Multilingual MMLU
49.3%स्वयं
Language
BoolQ
81.2%स्वयं
BIG-Bench Hard
70.4%स्वयं
Winogrande
67.0%स्वयं
Math
MATH-500
94.6%स्वयं
GSM8k
88.6%स्वयं
MATH
64.0%स्वयं
MGSM
63.9%स्वयं
AIMEMAA
57.5%स्वयं
Reasoning
HellaSwagAI2 (2019)
69.1%स्वयं
AA मूल्यांकन सूचकांक
(Artificial Analysis)Math Index(Artificial Analysis)6.7
Intelligence Index(Artificial Analysis)5.7
Coding Index(Artificial Analysis)3.8
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))0.7
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.5
Gpqa(NYU + Cohere + Anthropic (2023))0.3
Ifbench(Google Research (2023))0.2
Lcr(Artificial Analysis)0.2
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.1
Scicode(UIUC + Argonne National Lab (2024))0.1
Tau2(Sierra + U Toronto + Vector Institute (2025))0.1
Aime 25(MAA (Mathematical Association of America))0.1
Hle(Center for AI Safety + Scale AI (2025))0.0
Aime(MAA (Mathematical Association of America))0.0
Terminalbench V2 10.0
Terminalbench Hard(Stanford × Laude Institute (2026))0.0
LLM Stats श्रेणी स्कोर
(LLM Stats (zeroeval))Math70
Psychology70
Reasoning70
General70
Language60
Legal60
Finance60
Healthcare60
Physics50
Creativity50
Biology30
Chemistry30
Writing30
मूल्य निर्धारण
इनपुट मूल्यमुफ्त
आउटपुट मूल्यमुफ्त
मिश्रित मूल्य (3:1)मुफ्त
गति
टोकन/सेकंड45.1
पहले टोकन में देरी0.35s
पहले उत्तर में देरी0.35s
प्रदाता मूल्य रैंकिंग
प्रदाता मूल्य रैंकिंग
2 प्रदाता
सबसे सस्ता: Azure Cognitive Servicesसबसे महंगा: Azure
प्रदाताइनपुटआउटपुट
1Azure Cognitive Servicesसबसे सस्ता
$0.075
$0.3
2Azure
$0.075
$0.3
इस मॉडल के लिए विभिन्न API प्रदाताओं के मूल्य निर्धारण की तुलना करें।