Phi-4
Frontier SLM
Microsoft Research's 14B reasoning small model. Punches well above its weight on math and code, and runs on a single GPU.
Phi-4 is a reasoning AI model from Microsoft. It costs $0.070 per million input tokens and $0.280 per million output tokens (blended $0.133/M), with a 16,000-token context window.
INPUT
$0.070/M
per million input tokens
OUTPUT
$0.280/M
per million output tokens
BLENDED 70/30
$0.133/M
0.0%over 64 days · hover to read
CONTEXT
16,000
tokens
What it is good at
- 14B fits one GPU
- Strong math/code for size
- Permissive license
Typical use cases
- On-prem reasoning
- Edge inference
- Distillation studies
Benchmarks
vs. best public score
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.
More from Microsoft
See all 8 →Phi-4 Mini
Tiny · 128,000 ctx
in $0.040/Mout $0.160/M
Phi-3.5 Mini
Compact · 128,000 ctx
in $0.130/Mout $0.520/M
Phi-3.5 MoE
MoE · 128,000 ctx
in $0.160/Mout $0.640/M
Phi-3 Medium
Balanced · 128,000 ctx
in $0.170/Mout $0.680/M
Phi-3 Mini
Tiny · 128,000 ctx
in $0.130/Mout $0.520/M
Phi-3 Small
Compact · 128,000 ctx
in $0.150/Mout $0.600/M
Frequently asked questions
How much does Phi-4 cost?
Phi-4 costs $0.070 per million input tokens and $0.280 per million output tokens, for a blended reference rate of $0.133 per million tokens.
What is Phi-4's context window?
Phi-4 supports up to 16,000 tokens of context in a single request.
What is Phi-4 best for?
Phi-4 is well suited to 14B fits one GPU, Strong math/code for size and Permissive license.
Who makes Phi-4?
Phi-4 is developed and served by Microsoft.