Alibaba
The Qwen3.6 native vision-language Plus series models demonstrate exceptional performance on par with the current state-of-the-art models, with a significant improvement in overall results compared to the 3.5 series. The models have been markedly enhanced in code-related capabilities such as agentic coding, front-end programming, and Vibe coding, as well as in multi-modal general object recognition, OCR, and object localization.
Qwen model releasesRank #78 across 600
Rank #87 across 266
Rank #202 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 39.6 | #78 |
| Artificial Analysis Coding Index | coding | 54.5 | #87 |
| LiveBench Mathematics | math | 83.7% | #24 |
| GPQA | reasoning | 88.2% | #62 |
| Humanity's Last Exam |
| reasoning |
| 25.7% |
| #86 |
| SciCode | coding, reasoning | 40.7% | #138 |
| Output Speed | speed | 55.3 tok/s | #135 |
| Time to First Token | speed | 1.55s | #99 |
| Blended Price | cost | $1.13/M | #202 |
| Input Price | cost | $0.500/M | #187 |
| Output Price | cost | $3.00/M | #221 |
| Value Index | cost, overall | 35.2 | #115 |