Phi-4
MicrosoftPhiओपन वेटMIT · व्यावसायिक उपयोग
विवरण
phi-4 is a state-of-the-art open model built to excel at advanced reasoning, coding, and knowledge tasks. It leverages a blend of synthetic data, filtered web data, academic texts, and supervised fine-tuning for precision, alignment, and safety.
रिलीज़ तिथि
2024-12-12
पैरामीटर
14.7B
संदर्भ लंबाई
16K
मोडैलिटीज़
text
क्षमता रडार
25
general
23
coding
30
reasoning
41
science
28
agents
0
multimodal
रैंकिंग
| डोमेन | #रैंक | स्कोर | स्रोत |
|---|---|---|---|
| कोडिंग रैंकिंग | 586 | 10.0 | AA |
| सामान्य रैंकिंग | 569 | 21.0 | AA |
| विज्ञान | 475 | 30.0 | AA |
बेंचमार्क स्कोर (LLM Stats)
(LLM Stats (zeroeval))Chat
IFEvalGoogle Research (2023)
83.4%स्वयं
Factuality
SimpleQA
3.0%स्वयं
General
MMLU
84.8%स्वयं
Arena Hard
73.3%स्वयं
Language
MMLU-Pro
74.3%स्वयं
Math
MGSM
80.6%स्वयं
MATH
80.4%स्वयं
OmniMath
76.6%स्वयं
AIME 2024
75.3%स्वयं
AIME 2025
62.9%स्वयं
LiveBench
47.6%स्वयं
Reasoning
FlenQA
97.7%स्वयं
HumanEval+
92.9%स्वयं
HumanEvalOpenAI (2021)
82.6%स्वयं
DROP
75.5%स्वयं
PhiBench
70.6%स्वयं
GPQANYU + Cohere + Anthropic (2023)
65.8%स्वयं
LiveCodeBench
53.8%स्वयं
AA मूल्यांकन सूचकांक
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))81.0
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))71.4
Gpqa(NYU + Cohere + Anthropic (2023))57.5
Ifbench(Google Research (2023))23.5
Livecodebench(UC Berkeley + MIT + Cornell (2024))23.1
Math Index(Artificial Analysis)18.0
Aime 25(MAA (Mathematical Association of America))18.0
Aime(MAA (Mathematical Association of America))14.3
Intelligence Index(Artificial Analysis)5.9
Hle(Center for AI Safety + Scale AI (2025))3.8
Terminalbench Hard(Stanford × Laude Institute (2026))3.8
Lcr(Artificial Analysis)0.0
Tau2(Sierra + U Toronto + Vector Institute (2025))0.0
LLM Stats श्रेणी स्कोर
(LLM Stats (zeroeval))Language80
Legal80
Finance80
Healthcare80
Code80
Creativity80
Writing80
Chat70
Math70
Reasoning70
General70
Instruction Following60
Physics60
Structured Output60
Biology60
Chemistry60
Factuality0
मूल्य निर्धारण
इनपुट मूल्य$0.125 / 1M टोकन
आउटपुट मूल्य$0.5 / 1M टोकन
मिश्रित मूल्य (3:1)$0.219 / 1M टोकन
गति
टोकन/सेकंड43.9
पहले टोकन में देरी0.98s
पहले उत्तर में देरी0.98s
प्रदाता मूल्य रैंकिंग
प्रदाता मूल्य रैंकिंग
6 प्रदाता
सबसे सस्ता: DeepInfraसबसे महंगा: Azure
प्रदाताइनपुटआउटपुट
1DeepInfraसबसे सस्ता
$0
$0
2OpenRouter
$0.07
$0.14
3Kilo Gateway
$0.07
$0.14
4Microsoftप्राथमिक
$0.125
$0.5
5Azure Cognitive Services
$0.125
$0.5
6Azure
$0.125
$0.5
इस मॉडल के लिए विभिन्न API प्रदाताओं के मूल्य निर्धारण की तुलना करें।