Skip to content
Easy Benchmarks
Best ModelsToolsProvidersUpdatesOpen Full Benchmarks
EnglishFrançaisEspañolDeutschPortuguês

Easy Benchmarks

Tell us what you want to achieve. Easy Benchmarks turns current quality, price, and speed data into a simple shortlist.

Benchmark data is attributed to Artificial Analysis. Model metadata and catalog pricing may also use Vercel AI Gateway data when available.

Latest ReportBenchmark GlossaryMethodologyllms.txt

Best Models

Best AI Model for Time to First Token

Command A+ is currently the highest-ranked model for Time to First Token in the available data.

Open Full BenchmarksAI Model Finder
Current Leader
Command A+
Value
0.16 s
Dataset Coverage
314
Updated
Jul 16, 2026

How to use this result

  1. 1

    Start with the leader

    Use the top-ranked model as a starting point, not an automatic decision.

  2. 2

    Check the best alternatives

    Compare the top choices on price, speed, availability, and your own constraints.

  3. 3

    Verify the source

    Open the linked result page and methodology before quoting the ranking.

Top Models

Median time to first token. Lower means the model starts responding sooner, even if total generation speed differs.

RankModelProviderValue
#1Command A+Cohere0.16 s
#2North Mini CodeCohere0.18 s
#3NVIDIA Nemotron Nano 12B v2 VL (Reasoning)NVIDIA0.22 s
#4Tiny Aya GlobalCohere0.22 s
#5Llama Nemotron Super 49B v1.5 (Reasoning)NVIDIA

Data & Evidence

This answer uses the exact same metric and live database ranking as the linked Easy Benchmarks leaderboard.

Source: Artificial Analysis Time to First Token. Updated: Jul 16, 2026.

View the live benchmark resultsView the original result pageRead the benchmark methodology

Related Pages

Best AI Model for OverallBest AI Model for CodingBest AI Model for MathBest AI Model for Reasoning & KnowledgeBest AI Model for SpeedBest AI Model for Price & Value
0.26 s
#6Llama Nemotron Super 49B v1.5 (Non-reasoning)NVIDIA0.26 s
#7Gemini 2.5 Flash-Lite (Non-reasoning)Google0.32 s
#8Phi-4 Mini InstructMicrosoft0.32 s
#9Hermes 3 - Llama-3.1 70BNous Research0.34 s
#10Command ACohere0.34 s
#11Cogito v2.1 (Reasoning)Deep Cogito0.35 s
#12Phi-4 Multimodal InstructMicrosoft0.37 s
#13Ministral 3 3BMistral0.37 s
#14Gemma 3n E4B InstructGoogle0.38 s
#15Grok Build 0.1 0616SpaceXAI0.4 s
#16Qwen3.5 4B (Reasoning)Alibaba0.4 s
#17Mistral Small 3.2Mistral0.4 s
#18NVIDIA Nemotron 3 Nano 30B A3B (Non-reasoning)NVIDIA0.41 s
#19gpt-oss-20b (high)OpenAI0.41 s
#20Qwen3.5 4B (Non-reasoning)Alibaba0.43 s