DeepSeek
DeepSeek-V4 series incorporate several key upgrades in architecture and optimization: (1) a hybrid attention architecture that combines Compressed Sparse Attention (CSA) and Heavily Compressed Attention (HCA) to improve long-context efficiency; (2) ManifoldConstrained Hyper-Connections (mHC) that enhance conventional residual connections; (3) and the Muon optimizer for faster convergence and greater training stability
DeepSeek release notesRank #94 across 600
Rank #96 across 266
Rank #48 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
Streaming speed is not measured for this model yet.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 37.5 | #94 |
| Artificial Analysis Coding Index | coding | 52.0 | #96 |
| GPQA | reasoning | 86.7% | #77 |
| Humanity's Last Exam | reasoning | 27.8% | #74 |
| SciCode |
| coding, reasoning |
| 42.0% |
| #118 |
| Blended Price | cost | $0.175/M | #48 |
| Input Price | cost | $0.140/M | #58 |
| Output Price | cost | $0.280/M | #35 |
| Value Index | cost, overall | 214.3 | #10 |