Skip to content
Easy Benchmarks
Best ModelsToolsProvidersUpdatesOpen Full Benchmarks
EnglishFrançaisEspañolDeutschPortuguês

Easy Benchmarks

Tell us what you want to achieve. Easy Benchmarks turns current quality, price, and speed data into a simple shortlist.

Benchmark data is attributed to Artificial Analysis. Model metadata and catalog pricing may also use Vercel AI Gateway data when available.

Latest ReportBenchmark GlossaryMethodologyllms.txt

Best Models

Best AI Model for AIME

GPT-5 (high) is currently the highest-ranked model for AIME in the available data.

Open Full BenchmarksAI Model Finder
Current Leader
GPT-5 (high)
Value
95.7%
Dataset Coverage
194
Updated
Jul 16, 2026

How to use this result

  1. 1

    Start with the leader

    Use the top-ranked model as a starting point, not an automatic decision.

  2. 2

    Check the best alternatives

    Compare the top choices on price, speed, availability, and your own constraints.

  3. 3

    Verify the source

    Open the linked result page and methodology before quoting the ranking.

Top Models

AIME measures advanced competition math performance. Higher percentages indicate more solved problems.

RankModelProviderValue
#1GPT-5 (high)OpenAI95.7%
#2Grok 4SpaceXAI94.3%
#3o4-mini (high)OpenAI94%
#4Qwen3 235B A22B 2507 (Reasoning)Alibaba94%
#5Grok 3 mini Reasoning (high)SpaceXAI93.3%

Data & Evidence

This answer uses the exact same metric and live database ranking as the linked Easy Benchmarks leaderboard.

Source: Artificial Analysis AIME. Updated: Jul 16, 2026.

View the live benchmark resultsView the original result pageRead the benchmark methodology

Related Pages

Best AI Model for OverallBest AI Model for CodingBest AI Model for MathBest AI Model for Reasoning & KnowledgeBest AI Model for SpeedBest AI Model for Price & Value
#6
GPT-5 (medium)
OpenAI
91.7%
#7Qwen3 30B A3B 2507 (Reasoning)Alibaba90.7%
#8o3OpenAI90.3%
#9DeepSeek R1 0528 (May '25)DeepSeek89.3%
#10Gemini 2.5 ProGoogle88.7%
#11GLM-4.5 (Reasoning)Z AI87.3%
#12Gemini 2.5 Pro Preview (Mar' 25)Google87%
#13Llama Nemotron Super 49B v1.5 (Reasoning)NVIDIA86%
#14o3-mini (high)OpenAI86%
#15MiniMax M1 80kMiniMax84.7%
#16EXAONE 4.0 32B (Reasoning)LG AI Research84.3%
#17Gemini 2.5 Flash Preview (Reasoning)Google84.3%
#18Gemini 2.5 Pro Preview (May' 25)Google84.3%
#19Qwen3 235B A22B (Reasoning)Alibaba84%
#20GPT-5 (low)OpenAI83%