Alibaba
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...
Qwen model releasesRank #343 across 566
No rank
Rank #201 across 371
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 11.1 | #343 |
| Artificial Analysis Math Index | math | 68.3 | #100 |
| MMLU-Pro | reasoning | 79.1% | #136 |
| GPQA | reasoning | 67.1% | #281 |
| Humanity's Last Exam |
| reasoning |
| 6.3% |
| #287 |
| LiveCodeBench | coding | 51.4% | #146 |
| SciCode | coding, reasoning | 30.1% | #299 |
| Output Speed | speed | 80.6 tok/s | #182 |
| Time to First Token | speed | 1.07s | #153 |
| Blended Price | cost | $1.23/M | #201 |
| Input Price | cost | $0.700/M | #208 |
| Output Price | cost | $2.80/M | #202 |
| Value Index | cost, overall | 9.1 | #254 |