Alibaba
The Qwen3 series VL models has been comprehensively upgraded in areas such as visual coding and spatial perception. Its visual perception and recognition capabilities have significantly improved, supporting the understanding of ultra-long videos, and its OCR functionality has undergone a major enhancement.
Qwen model releasesRank #323 across 600
No rank
Rank #211 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
Streaming speed is not measured for this model yet.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 14.3 | #323 |
| Artificial Analysis Math Index | math | 70.7 | #95 |
| MMLU-Pro | reasoning | 82.3% | #64 |
| GPQA | reasoning | 71.2% | #257 |
| Humanity's Last Exam |
| reasoning |
| 6.3% |
| #302 |
| LiveCodeBench | coding | 59.4% | #117 |
| SciCode | coding, reasoning | 35.9% | #241 |
| Blended Price | cost | $1.23/M | #211 |
| Input Price | cost | $0.700/M | #221 |
| Output Price | cost | $2.80/M | #211 |
| Value Index | cost, overall | 11.7 | #235 |