DeepSeek
DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...
DeepSeek release notesRank #212 across 566
No rank
Rank #158 across 371
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
Streaming speed is not measured for this model yet.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 21.0 | #212 |
| Artificial Analysis Math Index | math | 49.7 | #141 |
| MMLU-Pro | reasoning | 83.3% | #60 |
| GPQA | reasoning | 73.5% | #220 |
| Humanity's Last Exam |
| reasoning |
| 6.3% |
| #285 |
| LiveCodeBench | coding | 57.7% | #124 |
| SciCode | coding, reasoning | 36.7% | #207 |
| Blended Price | cost | $0.840/M | #158 |
| Input Price | cost | $0.560/M | #179 |
| Output Price | cost | $1.68/M | #146 |
| Value Index | cost, overall | 25.0 | #134 |