Inception
A diffusion-based reasoning LLM that generates text via parallel refinement (not token-by-token), delivering real-time latency with ~1k tokens/sec plus 128K context and built-in tool/JSON support.
Inception model updatesRank #231 across 600
Rank #163 across 266
Rank #100 across 379
Percentile score by analysis domain.
* Cost is inverted: lower input, output, and blended prices rank higher.
Higher bars mean stronger relative placement.
| Metric | Domain | Value | Rank |
|---|---|---|---|
| Artificial Analysis Intelligence Index | overall | 21.4 | #231 |
| Artificial Analysis Coding Index | coding | 31.1 | #163 |
| GPQA | reasoning | 77.0% | #191 |
| Humanity's Last Exam | reasoning | 15.5% | #148 |
| SciCode |
| coding, reasoning |
| 38.7% |
| #182 |
| Output Speed | speed | 854.8 tok/s | #2 |
| Time to First Token | speed | 3.54s | #128 |
| Blended Price | cost | $0.375/M | #100 |
| Input Price | cost | $0.250/M | #116 |
| Output Price | cost | $0.750/M | #99 |
| Value Index | cost, overall | 57.1 | #83 |