Meta
Llama 4 Scout is the best multimodal model in the world in its class and is more powerful than our Llama 3 models, while fitting in a single H100 GPU. Additionally, Llama 4 Scout supports an industry-leading context window of up to 10M tokens.
Introducing Llama 4Rank #390 across 600
Rank #249 across 266
Rank #80 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 10.0 | #390 |
| Artificial Analysis Coding Index | coding | 8.2 | #249 |
| Artificial Analysis Math Index | math | 14.0 | #221 |
| MMLU-Pro | reasoning | 75.2% | #156 |
| reasoning |
| 58.7% |
| #357 |
| Humanity's Last Exam | reasoning | 4.3% | #448 |
| LiveCodeBench | coding | 29.9% | #222 |
| SciCode | coding, reasoning | 17.0% | #459 |
| MATH-500 | math | 84.4% | #98 |
| AIME | math | 28.3% | #88 |
| Output Speed | speed | 112.8 tok/s | #71 |
| Time to First Token | speed | 0.60s | #23 |
| Blended Price | cost | $0.300/M | #80 |
| Input Price | cost | $0.180/M | #75 |
| Output Price | cost | $0.660/M | #93 |
| Value Index | cost, overall | 33.3 | #121 |