Alibaba
Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique...
Qwen model releasesRank #442 across 566
No rank
Rank #93 across 371
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 6.8 | #442 |
| Artificial Analysis Math Index | math | 21.7 | #207 |
| MMLU-Pro | reasoning | 71.0% | #222 |
| GPQA | reasoning | 51.5% | #390 |
| Humanity's Last Exam |
| reasoning |
| 4.6% |
| #411 |
| LiveCodeBench | coding | 32.2% | #210 |
| SciCode | coding, reasoning | 26.4% | #362 |
| MATH-500 | math | 86.3% | #90 |
| AIME | math | 26.0% | #91 |
| Output Speed | speed | 113.9 tok/s | #120 |
| Time to First Token | speed | 0.98s | #128 |
| Blended Price | cost | $0.350/M | #93 |
| Input Price | cost | $0.200/M | #93 |
| Output Price | cost | $0.800/M | #102 |
| Value Index | cost, overall | 19.4 | #171 |