Alibaba
Qwen3-32B is a world-class model with comparable quality to DeepSeek R1 while outperforming GPT-4.1 and Claude Sonnet 3.7. It excels in code-gen, tool-calling, and advanced reasoning, making it an exceptional model for a wide range of production use cases.
Qwen model releasesRank #369 across 600
Rank #222 across 266
Rank #258 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
Streaming speed is not measured for this model yet.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 11.5 | #369 |
| Artificial Analysis Coding Index | coding | 15.3 | #222 |
| Artificial Analysis Math Index | math | 73.0 | #88 |
| GPQA | reasoning | 66.8% | #297 |
| reasoning |
| 8.3% |
| #245 |
| LiveCodeBench | coding | 54.6% | #133 |
| SciCode | coding, reasoning | 35.4% | #252 |
| MATH-500 | math | 96.1% | #35 |
| AIME | math | 80.7% | #23 |
| Blended Price | cost | $2.63/M | #258 |
| Input Price | cost | $0.700/M | #219 |
| Output Price | cost | $8.40/M | #277 |
| Value Index | cost, overall | 4.4 | #331 |