Modelle / Anthropic

Claude Opus 4 (2024-07)

Claude Opus · Release 2024-07-01

Claude-Opus-4-Checkpoint Juli 2024. HELM Safety 92,1 % im öffentlichen Ledger. Vorgänger von Opus 4.5+.

Früher Opus-4-API-Snapshot für Leaderboard-Kontinuität.

Composite88
Kontext200K
Input / 1M$15
Output / 1M$75
Wissensstichtag2024-04-01
Statuslightweight
Not enough families covered yet.Add more benchmark rows for Claude Opus 4 (2024-07) to see the radar.
Benchmark-Platzierungen

Wo Claude Opus 4 (2024-07) auf jeder öffentlichen Benchmark-Quelle mit Zeile steht. Rang zählt jedes Modell mit neuester Zeile im selben Test — kein universeller Qualitätsscore.

BenchmarkFamilyRangScoreQuelleDatum
HELM SafetySafety#1/ 992Stanford HELM leaderboard2024-08-15
MATHMath#9/ 2086Anthropic model card2025-06-01
Family-Kontext

Die drei höchstscorierenden kartierten Modelle pro Capability-Family. Wo Claude Opus 4 (2024-07) auftaucht, ist es hervorgehoben.

Family-Abdeckung

Ein hoher Composite, der eine schwache Family versteckt, ist eine Falle. Diese Balken zeigen Families ohne öffentlichen Test — und wo das Modell führt.

Math86
Safety92
Preis vs. Performance
0255075100$0.10$0.30$1$3$10$30input USD / 1M tokens · log scale →← intelligence

Hover or tab any dot for name, score, and input price.

AnthropicOpenAIMistralAlibabaDeepSeekGoogleMoonshotMiniMaxxAIMetaZhipuXiaomiMeituanCohere

Volles Benchmark-Ledger

Jeder Katalog-Benchmark für Claude Opus 4 (2024-07). Scores verlinken zur Originalquelle; Lücken heißen: noch keine öffentliche Zeile.

2 Quell-Zeilen

2 of 32 catalog benchmarks have a sourced row for Claude Opus 4 (2024-07).

Math

Safety

Quellen