InclusionAI
Ling 3.0 Flash is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers to complete more useful work within constrained token, latency, and serving-cost budgets.
InclusionAI model releasesRank #164 across 687
Rank #146 across 330
Rank #29 across 432
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 24.9 | #164 |
| Artificial Analysis Coding Index | coding | 50.6 | #146 |
| GPQA | reasoning | 85.5% | #133 |
| Humanity's Last Exam | reasoning | 23.7% | #169 |
| SciCode |
| coding, reasoning |
| 42.0% |
| #139 |
| Output Speed | speed | 362.8 tok/s | #3 |
| Time to First Token | speed | 1.73s | #102 |
| Blended Price | cost | $0.111/M | #29 |
| Input Price | cost | $0.075/M | #33 |
| Output Price | cost | $0.220/M | #30 |
| Value Index | cost, overall | 224.3 | #5 |