DeepSeek
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team.
DeepSeek release notesRank #324 across 600
Rank #189 across 266
Rank #123 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
Streaming speed is not measured for this model yet.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 14.2 | #324 |
| Artificial Analysis Coding Index | coding | 23.0 | #189 |
| Artificial Analysis Math Index | math | 26.0 | #195 |
| MMLU-Pro | reasoning | 75.2% | #153 |
| reasoning |
| 55.7% |
| #377 |
| Humanity's Last Exam | reasoning | 3.6% | #521 |
| LiveCodeBench | coding | 35.9% | #194 |
| SciCode | coding, reasoning | 35.4% | #249 |
| MATH-500 | math | 88.7% | #80 |
| AIME | math | 25.3% | #92 |
| Blended Price | cost | $0.493/M | #123 |
| Input Price | cost | $0.360/M | #158 |
| Output Price | cost | $0.890/M | #114 |
| Value Index | cost, overall | 28.8 | #133 |