Models / Anthropic

Claude Opus 4 (2024-07)

Claude Opus · Released 2024-07-01

July 2024 Claude Opus 4 checkpoint. HELM safety 92.1% on the public ledger. Predecessor to Opus 4.5+.

Early Opus 4 API snapshot for leaderboard continuity.

Composite88
Context200K
Input / 1M$15
Output / 1M$75
Knowledge cutoff2024-04-01
Statuslightweight
Not enough families covered yet.Add more benchmark rows for Claude Opus 4 (2024-07) to see the radar.
Benchmark placements

Where Claude Opus 4 (2024-07) places on each public benchmark source that publishes a row. Rank counts every model with a latest row in the same test — not a universal quality score.

BenchmarkFamilyRankScoreSourceDate
HELM SafetySafety#1/ 992Stanford HELM leaderboard2024-08-15
MATHMath#9/ 2086Anthropic model card2025-06-01
Family context

The three highest-scoring carded models in each capability family. Where Claude Opus 4 (2024-07) shows up, it's highlighted.

Family coverage

A high composite that hides a weak family is a trap. These bars surface the families where this model hasn't been publicly tested, and where it leads.

Math86
Safety92
Price vs. performance
0255075100$0.10$0.30$1$3$10$30input USD / 1M tokens · log scale →← intelligence

Hover or tab any dot for name, score, and input price.

AnthropicOpenAIMistralAlibabaDeepSeekGoogleMoonshotMiniMaxxAIMetaZhipuXiaomiMeituanCohere

Full benchmark ledger

Every catalog benchmark for Claude Opus 4 (2024-07). Scores link to the original source; gaps mean no public row exists yet.

2 sourced rows

2 of 32 catalog benchmarks have a sourced row for Claude Opus 4 (2024-07).

Math

Safety

Sources