Zum Inhalt springen
Easy Benchmarks
Beste ModelleToolsAnbieterUpdatesVollständige Benchmarks öffnen
EnglishFrançaisEspañolDeutschPortuguês

Easy Benchmarks

Sag uns, was du erreichen möchtest. Easy Benchmarks übersetzt aktuelle Benchmark-, Preis- und Geschwindigkeitsdaten in eine einfache Auswahl.

Benchmark data is attributed to Artificial Analysis. Model metadata and catalog pricing may also use Vercel AI Gateway data when available.

Aktueller BerichtBenchmark-GlossarMethodikllms.txt

Beste Modelle

Bestes KI-Modell für Geschwindigkeit

Mercury 2 ist derzeit das bestplatzierte Modell für Geschwindigkeit in den verfügbaren Daten.

Vollständige Benchmarks öffnenKI-Modellfinder
Aktueller Spitzenreiter
Mercury 2
Wert
880,1 tok/s
Datensatzabdeckung
314
Aktualisiert
16.07.2026

So nutzt du dieses Ergebnis

  1. 1

    Mit dem Spitzenreiter beginnen

    Nutze den Erstplatzierten als Ausgangspunkt, nicht als automatische Entscheidung.

  2. 2

    Die besten Optionen prüfen

    Vergleiche die ersten Modelle bei Preis, Geschwindigkeit und Verfügbarkeit.

  3. 3

    Die Quelle prüfen

    Öffne die verlinkte Ergebnisseite und Methodik, bevor du das Ergebnis zitierst.

Top-Modelle

Median generated output tokens per second. Higher means the model streams completions faster after it starts responding.

RangModellAnbieterWert
#1Mercury 2Inception880,1 tok/s
#2Granite 4.0 H SmallIBM439,5 tok/s
#3Granite 3.3 8B (Non-reasoning)IBM414,9 tok/s
#4LFM2.5-VL-1.6BLiquid AI392,7 tok/s
#5Step 3.7 FlashStepFun385,2 tok/s

Daten & Nachweise

This answer uses the exact same metric and live database ranking as the linked Easy Benchmarks leaderboard.

Quelle: Artificial Analysis Output Speed. Aktualisiert: 16.07.2026.

View the live benchmark resultsView the original result pageRead the benchmark methodology

Verwandte Seiten

Bestes KI-Modell für allgemeine LeistungBestes KI-Modell für ProgrammierungBestes KI-Modell für MathematikBestes KI-Modell für Schlussfolgern und WissenBestes KI-Modell für Preis-LeistungBestes KI-Modell für Artificial Analysis Math Index
#6HyperNova 60B 2605Multiverse Computing350,9 tok/s
#7LFM2.5-8B-A1BLiquid AI343,4 tok/s
#8Nemotron 3 Nano Omni 30B A3B ReasoningNVIDIA320 tok/s
#9gpt-oss-120b (low)OpenAI296,3 tok/s
#10Gemini 3.1 Flash-LiteGoogle291,6 tok/s
#11Llama 3.1 Nemotron Instruct 70BNVIDIA286,3 tok/s
#12NVIDIA Nemotron Nano 12B v2 VL (Reasoning)NVIDIA282,5 tok/s
#13gpt-oss-20b (low)OpenAI260,9 tok/s
#14Step 3.5 Flash 2603StepFun256,9 tok/s
#15Nova MicroAmazon249,8 tok/s
#16Step 3.5 FlashStepFun247,4 tok/s
#17GPT-5.6 Luna (max)OpenAI243,6 tok/s
#18Gemini 3.5 Flash (high)Google241,2 tok/s
#19Gemini 2.5 Flash-Lite (Reasoning)Google237,8 tok/s
#20Gemini 3.5 Flash (medium)Google234,7 tok/s