Kimi
Kimi K2 Thinking is an advanced open-source thinking model by Moonshot AI. It can execute up to 200 – 300 sequential tool calls without human interference, reasoning coherently across hundreds of steps to solve complex problems. Built as a thinking agent, it reasons step by step while using tools, achieving state-of-the-art performance on Humanity's Last Exam (HLE), BrowseComp, and other benchmarks, with major gains in reasoning, agentic search, coding, writing, and general capabilities.
Kimi model releasesRank #135 across 600
Rank #200 across 266
Rank #197 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
Streaming speed is not measured for this model yet.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 32.7 | #135 |
| Artificial Analysis Coding Index | coding | 21.0 | #200 |
| Artificial Analysis Math Index | math | 94.7 | #11 |
| MMLU-Pro | reasoning | 84.8% | #34 |
| reasoning |
| 83.8% |
| #118 |
| Humanity's Last Exam | reasoning | 22.3% | #105 |
| LiveCodeBench | coding | 85.3% | #14 |
| SciCode | coding, reasoning | 42.4% | #114 |
| Blended Price | cost | $1.08/M | #197 |
| Input Price | cost | $0.600/M | #200 |
| Output Price | cost | $2.50/M | #194 |
| Value Index | cost, overall | 30.4 | #128 |