DeepSeek
DeepSeek-V3.1-Terminus delivers more stable & reliable outputs across benchmarks compared to the previous version and addresses user feedback (i.e. language consistency and agent upgrades).
DeepSeek release notesRank #154 across 600
Rank #126 across 266
Rank #239 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
Streaming speed is not measured for this model yet.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 30.4 | #154 |
| Artificial Analysis Coding Index | coding | 43.5 | #126 |
| Artificial Analysis Math Index | math | 89.7 | #27 |
| MMLU-Pro | reasoning | 85.1% | #30 |
| reasoning |
| 79.2% |
| #163 |
| Humanity's Last Exam | reasoning | 15.2% | #150 |
| LiveCodeBench | coding | 79.8% | #29 |
| SciCode | coding, reasoning | 40.6% | #139 |
| Blended Price | cost | $1.91/M | #239 |
| Input Price | cost | $1.64/M | #279 |
| Output Price | cost | $2.75/M | #206 |
| Value Index | cost, overall | 15.9 | #200 |