DeepSeek
DeepSeek-V3.1-Terminus delivers more stable & reliable outputs across benchmarks compared to the previous version and addresses user feedback (i.e. language consistency and agent upgrades).
DeepSeek release notesRank #230 across 600
Rank #125 across 266
Rank #116 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
Streaming speed is not measured for this model yet.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 21.4 | #230 |
| Artificial Analysis Coding Index | coding | 43.5 | #125 |
| Artificial Analysis Math Index | math | 53.7 | #134 |
| MMLU-Pro | reasoning | 83.6% | #48 |
| reasoning |
| 75.1% |
| #217 |
| Humanity's Last Exam | reasoning | 8.4% | #243 |
| LiveCodeBench | coding | 52.9% | #139 |
| SciCode | coding, reasoning | 32.1% | #294 |
| Blended Price | cost | $0.453/M | #116 |
| Input Price | cost | $0.270/M | #123 |
| Output Price | cost | $1.00/M | #121 |
| Value Index | cost, overall | 47.2 | #94 |