Microsoft
[Microsoft Research](/microsoft) Phi-4 is designed to perform well in complex reasoning tasks and can operate efficiently in situations with limited memory or where quick responses are needed. At 14 billion...
Azure AI Foundry modelsRank #513 across 600
No rank
Rank #63 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 4.9 | #513 |
| Artificial Analysis Math Index | math | 18.0 | #214 |
| MMLU-Pro | reasoning | 71.4% | #185 |
| GPQA | reasoning | 57.5% | #365 |
| Humanity's Last Exam |
| reasoning |
| 4.1% |
| #478 |
| LiveCodeBench | coding | 23.1% | #263 |
| SciCode | coding, reasoning | 26.0% | #378 |
| MATH-500 | math | 81.0% | #106 |
| AIME | math | 14.3% | #114 |
| Output Speed | speed | 41.9 tok/s | #150 |
| Time to First Token | speed | 0.47s | #14 |
| Blended Price | cost | $0.219/M | #63 |
| Input Price | cost | $0.125/M | #50 |
| Output Price | cost | $0.500/M | #77 |
| Value Index | cost, overall | 22.4 | #164 |