Best Models
Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) is currently the highest-ranked model for SciCode in the available data.
Use the top-ranked model as a starting point, not an automatic decision.
Compare the top choices on price, speed, availability, and your own constraints.
Open the linked result page and methodology before quoting the ranking.
SciCode focuses on scientific programming tasks, so it contributes to both coding and reasoning comparisons.
| Rank | Model | Provider | Value |
|---|---|---|---|
| #1 | Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) | Anthropic | 60.2% |
| #2 | Gemini 3.1 Pro Preview | 58.9% | |
| #3 | Muse Spark 1.1 (xhigh) | Meta | 58.2% |
| #4 | GPT-5.6 Sol (high) | OpenAI | 56.9% |
| #5 | GPT-5.4 (xhigh) | OpenAI | 56.6% |
This answer uses the exact same metric and live database ranking as the linked Easy Benchmarks leaderboard.
Source: Artificial Analysis SciCode. Updated: Jul 16, 2026.
| #6 | GPT-5.6 Sol (medium) | OpenAI | 56.5% |
| #7 | GPT-5.6 Sol (max) | OpenAI | 56.1% |
| #8 | GPT-5.5 (xhigh) | OpenAI | 56.1% |
| #9 | Gemini 3 Pro Preview (high) | 56.1% |
| #10 | GPT-5.6 Sol (xhigh) | OpenAI | 56% |
| #11 | GPT-5.5 (high) | OpenAI | 55.9% |
| #12 | GPT-5.6 Sol (low) | OpenAI | 55.4% |
| #13 | GPT-5.2 Codex (xhigh) | OpenAI | 54.6% |
| #14 | Claude Opus 4.7 (Adaptive Reasoning, Max Effort) | Anthropic | 54.5% |
| #15 | Grok 4.5 (high) | SpaceXAI | 54.1% |
| #16 | GPT-5.6 Terra (max) | OpenAI | 53.9% |
| #17 | Claude Sonnet 5 (Adaptive Reasoning, Max Effort) | Anthropic | 53.6% |
| #18 | GPT-5.5 (medium) | OpenAI | 53.5% |
| #19 | Claude Opus 4.8 (Adaptive Reasoning, Max Effort) | Anthropic | 53.5% |
| #20 | Kimi K2.6 | Kimi | 53.5% |