This model improves upon Gemini 2.5 Pro and is catered towards challenging tasks, especially those involving complex reasoning or agentic workflows. Improvements highlighted include use cases for coding, multi-step function calling, planning, reasoning, deep knowledge tasks, and instruction following.
Introducing Gemini 3Rank #41 across 600
Rank #40 across 266
Rank #308 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 46.5 | #41 |
| Artificial Analysis Coding Index | coding | 68.8 | #40 |
| MMLU-Pro | reasoning | 91.2% | #1 |
| LiveBench Mathematics | math | 91.0% | #9 |
| GPQA |
| reasoning |
| 94.1% |
| #1 |
| Humanity's Last Exam | reasoning | 44.7% | #9 |
| SciCode | coding, reasoning | 58.9% | #2 |
| Output Speed | speed | 122 tok/s | #66 |
| Time to First Token | speed | 17.34s | #154 |
| Blended Price | cost | $4.50/M | #308 |
| Input Price | cost | $2.00/M | #296 |
| Output Price | cost | $12.00/M | #308 |
| Value Index | cost, overall | 10.3 | #246 |