Models / Anthropic

Claude Opus 4.5 (Reasoning)

Claude Opus · Released 2025-11-24

AA Intelligence Index 40.8. Listed on Artificial Analysis with API pricing $5/$25 per 1M tokens. Benchmark rows sync from the AA snapshot — see the ledger on this page.

Frontier model tracked on Artificial Analysis. Use the benchmark ledger below for sourced scores; we do not infer numbers beyond published rows.

#48 of 168 on AA Index · current snapshot
42 AA Index · 12 benchmark rows · 1 sourcesPosition is benchmark-specific — not a cross-family or cross-source ranking.
Composite64
Context128K
Input / 1M$5.0
Output / 1M$25
Knowledge cutoffopen
Statussolid

Family profile

Best published score in each covered benchmark family.

80/100 avg
Knowledge
90
Reasoning
87
Math
91
Coding
87
Agentic
47
Long context
76

6 tested benchmark families

Benchmark placements

Where Claude Opus 4.5 (Reasoning) places on each public benchmark source that publishes a row. Rank counts every model with a latest row in the same test — not a universal quality score.

BenchmarkFamilyRankScoreSourceDate
MMLU-ProKnowledge#2/ 3990Artificial Analysis2026-09-01
LiveCodeBenchCoding#6/ 3187Artificial Analysis2026-09-01
AIMEMath#17/ 3791Artificial Analysis2026-09-01
AA-LCRLong context#42/ 17976Artificial Analysis2026-09-01
SciCodeCoding#49/ 18050Artificial Analysis2026-09-01
Artificial Analysis Intelligence IndexReasoning#61/ 18742Artificial Analysis2026-09-01
Humanity's Last ExamReasoning#73/ 18030Artificial Analysis2026-09-01
GPQA DiamondReasoning#79/ 18587Artificial Analysis2026-09-01
IFBenchReasoning#83/ 12758Artificial Analysis2026-09-01
Terminal-BenchAgentic#92/ 18447Artificial Analysis2026-09-01
AA time to first tokenPerformance#93/ 186Artificial Analysis2026-09-01
AA output speedPerformance#94/ 187Artificial Analysis2026-09-01
Family context

The three highest-scoring models with pages in each capability family. Where Claude Opus 4.5 (Reasoning) shows up, it's highlighted.

KnowledgeClaude Opus 4.5 (Reasoning) · 90
  1. 1
    91
  2. 2
    91
  3. 3
    91
ReasoningClaude Opus 4.5 (Reasoning) · 87
  1. 1
    95
  2. 2
    95
  3. 3
    95
MathClaude Opus 4.5 (Reasoning) · 91
  1. 1
    99
  2. 2
    99
  3. 3
    99
CodingClaude Opus 4.5 (Reasoning) · 87
  1. 1
    96
  2. 2
    96
  3. 3
    96
AgenticClaude Opus 4.5 (Reasoning) · 47
  1. 1
    100
  2. 2
    100
  3. 3
    100
Long contextClaude Opus 4.5 (Reasoning) · 76
  1. 1
    83
  2. 2
    83
  3. 3
    81
Newest receipts

Benchmark rows added to the public ledger for Claude Opus 4.5 (Reasoning) in the last 120 days. Older rows live in the full table below.

12 recent
Family coverage

A high composite that hides a weak family is a trap. These bars surface the families where this model hasn't been publicly tested, and where it leads.

Knowledge90
Reasoning87
Math91
Coding87
Agentic47
Long context76
Price vs. performance
0255075100$0.10$0.30$1$3$10$30input USD / 1M tokens · log scale →← intelligence

Hover any dot for name, score, and input price. Keyboard: tab through the top twelve, or use the ranking below.

AnthropicOpenAIMistralAlibabaDeepSeekGoogleMoonshotMetaMiniMaxxAIZhipuXiaomiOtherMeituanCohereNVIDIA

Full benchmark ledger

Every catalog benchmark for Claude Opus 4.5 (Reasoning). Scores link to the original source; gaps mean no public row exists yet.

12 sourced rows

12 of 37 catalog benchmarks have a sourced row for Claude Opus 4.5 (Reasoning).

Knowledge

Reasoning

Math

Coding

Agentic

Long context

Performance

Sources