DeepSeek
DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...
DeepSeek release notesRank #362 across 566
No rank
Rank #152 across 371
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 9.9 | #362 |
| Artificial Analysis Math Index | math | 53.7 | #133 |
| MMLU-Pro | reasoning | 79.5% | #128 |
| GPQA | reasoning | 40.2% | #450 |
| Humanity's Last Exam |
| reasoning |
| 6.1% |
| #296 |
| LiveCodeBench | coding | 26.6% | #248 |
| SciCode | coding, reasoning | 31.3% | #287 |
| MATH-500 | math | 93.5% | #52 |
| AIME | math | 67.0% | #46 |
| Output Speed | speed | 22.6 tok/s | #313 |
| Time to First Token | speed | 0.62s | #65 |
| Blended Price | cost | $0.787/M | #152 |
| Input Price | cost | $0.700/M | #201 |
| Output Price | cost | $1.05/M | #119 |
| Value Index | cost, overall | 12.6 | #214 |