Mistral
Complex thinking, backed by deep understanding, with transparent reasoning you can follow and verify. The model excels in maintaining high-fidelity reasoning across numerous languages, even when switching between languages mid-task.
Introducing Mistral 3Rank #345 across 600
No rank
No rank
Percentile score by analysis domain.
Higher bars mean stronger relative placement.
Streaming speed is not measured for this model yet.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 12.5 | #345 |
| Artificial Analysis Math Index | math | 40.3 | #158 |
| MMLU-Pro | reasoning | 75.3% | #152 |
| GPQA | reasoning | 67.9% | #287 |
| Humanity's Last Exam |
| reasoning |
| 9.5% |
| #227 |
| LiveCodeBench | coding | 52.7% | #140 |
| SciCode | coding, reasoning | 29.7% | #316 |
| MATH-500 | math | 91.7% | #66 |
| AIME | math | 70.0% | #40 |