Gemini 2.5 Flash-Lite is a balanced, low-latency model with configurable thinking budgets and tool connectivity (e.g., Google Search grounding and code execution). It supports multimodal input and offers a 1M-token context window.
Gemini 2.5 thinking modelsRank #468 across 600
No rank
Rank #51 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
Streaming speed is not measured for this model yet.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 6.9 | #468 |
| Artificial Analysis Math Index | math | 35.3 | #173 |
| GPQA | reasoning | 47.4% | #423 |
| Humanity's Last Exam | reasoning | 3.7% | #511 |
| LiveCodeBench |
| coding |
| 40.0% |
| #182 |
| SciCode | coding, reasoning | 17.7% | #453 |
| MATH-500 | math | 92.6% | #61 |
| AIME | math | 50.0% | #59 |
| Blended Price | cost | $0.175/M | #51 |
| Input Price | cost | $0.100/M | #33 |
| Output Price | cost | $0.400/M | #50 |
| Value Index | cost, overall | 39.4 | #107 |