Phi-4
MicrosoftPhiOpen WeightMIT · Commercial OK
Description
phi-4 is a state-of-the-art open model built to excel at advanced reasoning, coding, and knowledge tasks. It leverages a blend of synthetic data, filtered web data, academic texts, and supervised fine-tuning for precision, alignment, and safety.
Release Date
2024-12-12
Parameters
14.7B
Context Length
16K
Modalities
text
Capability Radar
28
general
17
coding
30
reasoning
36
scienceest.
0
agents
0
multimodal
Science uses a reasoning proxy when dedicated science benchmarks are unavailable.
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Code Ranking | 390 | 14.0 | AA |
| General Ranking | 415 | 23.0 | AA |
| Math Reasoning | 267 | 30.0 | AA |
| Reasoning | 25 | 83.0 | LS |
| Science | 295 | 36.0 | AA |
Benchmark Scores (LLM Stats)
Biology
GPQA
56.1%SR
Code
HumanEval
82.6%SR
Creativity
Arena Hard
75.4%SR
Factuality
SimpleQA
3.0%SR
Finance
MMLU
84.8%SR
MMLU-Pro
70.4%SR
General
IFEval
63.0%SR
PhiBench
56.2%SR
LiveBench
47.6%SR
Math
MGSM
80.6%SR
MATH
80.4%SR
DROP
75.5%SR
Reasoning
HumanEval+
82.8%SR
AA Evaluation Indices
Math Index18.0
Coding Index11.2
Intelligence Index10.4
Math 5000.8
Mmlu Pro0.7
Gpqa0.6
Scicode0.3
Ifbench0.2
Livecodebench0.2
Aime 250.2
Aime0.1
Hle0.0
Terminalbench Hard0.0
Lcr0.0
Tau20.0
LLM Stats Category Scores
Writing80
Code80
Creativity80
Finance80
Healthcare80
Language80
Legal80
Math70
Reasoning70
Structured Output60
Biology60
Chemistry60
General60
Instruction Following60
Physics60
Factuality0
Pricing
Input Price$0.125 / 1M tokens
Output Price$0.5 / 1M tokens
Blended Price (3:1)$0.219 / 1M tokens
Speed
Tokens/sec38.5 tokens/s
Time to First Token0.51s
Time to Answer0.51s
Available Providers
(LS internal units)No provider data available