Benchmarks / Performance

AA time to first token

TTFT

Median seconds until the first token returns on the model's default API provider with a ~1k-token input prompt.

Artificial AnalysisPerformanceFlagshipActiveUnknown contamination riskSince 2026
What this does not measure
  • End-to-end task time for agent loops — TTFT is one hop, not a full coding session.
  • Non-default providers — AA pins one default route per model for comparability.
Analysis

Why this benchmark is useful

Editorial brief pendingWe publish the methodology and ledger first; benchmark-specific analysis ships after desk review.

Scope

Coverage map

Task family
Performance
Format
Single-shot API request with streaming enabled; median P50 over 72 hours.
Scoring
Lower seconds is faster.
Maintainer
Artificial Analysis
Reading guide

How to read the scores

Reading guide pending. Use task format, scoring method, and source dates in the ledger until the desk brief ships.

Blind spots

What it does not cover

  • End-to-end task time for agent loops — TTFT is one hop, not a full coding session.
  • Non-default providers — AA pins one default route per model for comparability.
Scores

Evidence ledger

69 rows
6969 rows
129.43sbest score
69source-checked
1sources
2026-07-21source date
69

api · api

Distribution

Where the rows land

0255075100

Normalized to this benchmark's axis (0–100). Open the table for raw units.

Timeline

Newest receipts

  1. GPT 5.4129.43s ·
  2. GPT-5.6 Terra100.45s ·
  3. GPT-5.297.23s ·
  4. Claude Sonnet 588.49s ·
  5. GPT-5.6 Luna83.80s ·

Scores use this benchmark's own unit and axis, not a universal quality score.

#ModelRelease dateScoreProvenanceTrust
1GPT 5.4OpenAI2026-03-05129.43s
2026-07-21
Source-checked
2GPT-5.6 TerraOpenAI2026-07-09100.45s
2026-07-21
Source-checked
3GPT-5.2OpenAI2025-12-1197.23s
2026-07-21
Source-checked
4Claude Sonnet 5Anthropic2026-06-3088.49s
2026-07-21
Source-checked
5GPT-5.6 LunaOpenAI2026-07-0983.80s
2026-07-21
Source-checked
6GPT-5.6 SolOpenAI2026-07-0976.91s
2026-07-21
Source-checked
7Claude Fable 5Anthropic2026-06-0968.33s
2026-07-21
Source-checked
8Claude Sonnet 4.6 (Adaptive Reasoning, Max Effort)Anthropic2026-02-1757.98s
2026-07-21
Source-checked
9GPT-5.3 Codex (xhigh)OpenAI2026-02-0541.56s
2026-07-21
Source-checked
10GPT-5.1OpenAI2025-11-1338.96s
2026-07-21
Source-checked
11GPT 5.5OpenAI2026-04-2331.38s
2026-07-21
Source-checked
12Grok 4.3xAI2026-04-3031.32s
2026-07-21
Source-checked
13Claude Opus 4.8Anthropic2026-05-2830.13s
2026-07-21
Source-checked
14Gemini 3.1 Pro PreviewGoogle2026-02-1921.91s
2026-07-21
Source-checked
15Grok 4.20 ReasoningxAI2026-03-0516.81s
2026-07-21
Source-checked
16Claude Opus 4.7Anthropic2026-04-1616.72s
2026-07-21
Source-checked
17Claude Opus 4.6 (Adaptive Reasoning, Max Effort)Anthropic2026-02-0514.02s
2026-07-21
Source-checked
18GPT-5.5 (high)OpenAI2026-04-2313.95s
2026-07-21
Source-checked
19Gemini 3.5 FlashGoogle2026-05-1912.11s
2026-07-21
Source-checked
20Grok 4.5xAI2026-07-089.69s
2026-07-21
Source-checked
21Gemini 3.5 Flash (medium)Google2026-05-198.89s
2026-07-21
Source-checked
22GPT-5.4 MiniOpenAI2026-03-178.20s
2026-07-21
Source-checked
23GPT-5.5 (medium)OpenAI2026-04-236.88s
2026-07-21
Source-checked
24Kimi K3Moonshot2026-07-166.19s
2026-07-21
Source-checked
25Gemini 3.1 Flash LiteGoogle2026-05-074.90s
2026-07-21
Source-checked
26OpenAI o3OpenAI2025-04-164.84s
2026-07-21
Source-checked
27GPT-5.4 NanoOpenAI2026-03-173.46s
2026-07-21
Source-checked
28MiMo-V2.5-ProXiaomi2026-04-222.38s
2026-07-21
Source-checked
29Claude Opus 4.6Anthropic2026-02-051.85s
2026-07-21
Source-checked
30LFM2.5-8B-A1BLiquid2026-06-081.79s
2026-07-21
Source-checked
31Qwen 3.7 PlusAlibaba2026-06-021.69s
2026-07-21
Source-checked
32GPT-5.5 (low)OpenAI2026-04-231.61s
2026-07-21
Source-checked
33Claude Opus 4.7 (Non-reasoning, High Effort)Anthropic2026-04-161.55s
2026-07-21
Source-checked
34Qwen3.7 MaxAlibaba2026-05-191.55s
2026-07-21
Source-checked
35Gemma 4 12B (Non-reasoning)Google2026-06-101.49s
2026-07-21
Source-checked
36Gemma 4 12B (Reasoning)Google2026-06-071.46s
2026-07-21
Source-checked
37Claude Opus 4.5Anthropic2025-11-241.32s
2026-07-21
Source-checked
38MiniMax M3MiniMax2026-05-311.26s
2026-07-21
Source-checked
39Claude Sonnet 4.6Anthropic2026-02-171.18s
2026-07-21
Source-checked
40Kimi K2.7 CodeMoonshot2026-06-121.17s
2026-07-21
Source-checked
41MiniMax M2.7MiniMax2026-04-151.17s
2026-07-21
Source-checked
42Kimi K2.6Moonshot2026-04-201.08s
2026-07-21
Source-checked
43Kimi K2.5Moonshot2026-01-271.02s
2026-07-21
Source-checked
44MiniMax M2.5MiniMax2026-04-011.02s
2026-07-21
Source-checked
45DeepSeek V4 Pro (Reasoning, Max Effort)DeepSeek2026-04-241.01s
2026-07-21
Source-checked
46Gemma 4 31B ITGoogle2026-03-010.98s
2026-07-21
Source-checked
47Nemotron 3 Ultra 550B A55B (Reasoning)NVIDIA2026-06-040.88s
2026-07-21
Source-checked
48GLM-5.1Zhipu2026-04-070.86s
2026-07-21
Source-checked
49GLM-5.2Zhipu2026-06-130.84s
2026-07-21
Source-checked
50Muse Spark 1.1Meta2026-07-090.83s
2026-07-21
Source-checked
51Gemini 3 Flash PreviewGoogle2025-12-170.72s
2026-07-21
Source-checked
52Step 3.7 FlashStepFun2026-06-010.72s
2026-07-21
Source-checked
53HyperNova 60B 2605Multiverse2026-05-060.60s
2026-07-21
Source-checked
54Llama 4 MaverickMeta2026-04-050.60s
2026-07-21
Source-checked
55Mistral Large 3Mistral2025-12-020.59s
2026-07-21
Source-checked
56Llama 4 ScoutMeta2026-04-050.56s
2026-07-21
Source-checked
57Command ACohere2026-03-150.36s
2026-07-21
Source-checked
58North Mini CodeCohere2026-06-090.23s
2026-07-21
Source-checked
59Claude 3 OpusAnthropic2024-03-040.00s
2026-07-21
Source-checked
60DeepSeek Coder V2DeepSeek2024-05-150.00s
2026-07-21
Source-checked
61DeepSeek R1DeepSeek2025-01-200.00s
2026-07-21
Source-checked
62Gemini 1.0 UltraGoogle2024-02-080.00s
2026-07-21
Source-checked
63Gemini 3 ProGoogle2025-11-180.00s
2026-07-21
Source-checked
64GPT-5.4 ProOpenAI2026-03-050.00s
2026-07-21
Source-checked
65GPT-5.5 ProOpenAI2026-04-230.00s
2026-07-21
Source-checked
66LongCat-2.0Meituan2026-06-290.00s
2026-07-21
Source-checked
67MiniCPM5-1B (Reasoning)OpenBMB2026-06-040.00s
2026-07-21
Source-checked
68Muse SparkOther2026-01-010.00s
2026-07-21
Source-checked
69OpenAI o1OpenAI2024-09-120.00s
2026-07-21
Source-checked
Method

What it covers

Data-quality note pending. Every ledger row still carries source URL, source date, and ingest timestamp.

Score ceiling

Where it breaks down

No ceiling note is recorded yet. Treat clustering near the top as a warning that the benchmark may no longer separate frontier models.

Tools that report it

Receipts

Sources and further reading

The Pack · Editorial newsletter

New cards in your inbox. Free.

One short email when a card ships or changes status. No tracking, no third-party analytics. Unsubscribe in one click.