DeepSeek
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...
DeepSeek release notesRank #275 across 566
Rank #162 across 222
Rank #192 across 371
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
Streaming speed is not measured for this model yet.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 15.4 | #275 |
| Artificial Analysis Coding Index | coding | 21.2 | #162 |
| Artificial Analysis Math Index | math | 41.0 | #157 |
| MMLU-Pro | reasoning | 81.9% | #83 |
| reasoning |
| 65.5% |
| #300 |
| Humanity's Last Exam | reasoning | 5.2% | #338 |
| LiveCodeBench | coding | 40.5% | #178 |
| SciCode | coding, reasoning | 35.8% | #231 |
| MATH-500 | math | 94.2% | #47 |
| AIME | math | 52.0% | #54 |
| Blended Price | cost | $1.17/M | #192 |
| Input Price | cost | $1.14/M | #239 |
| Output Price | cost | $1.25/M | #135 |
| Value Index | cost, overall | 13.2 | #213 |