Modelle / Cohere

Command A

Command · Release 2026-03-15

Cohere's Enterprise-fokussiertes Command-Tier. Stark bei RAG, Tool-Use und mehrsprachigen Retrieval-Workflows. 256K-Kontext.

Wahl, wenn der Stack bereits Cohere Embed + Rerank nutzt und ein Anbieter für Retrieval-Pipelines reicht.

#48 von 51 im AA Index · aktueller Snapshot
8 AA Index · 14 Benchmark-Zeilen · 2 QuellenPosition ist benchmark-spezifisch — kein Cross-Family- oder Cross-Source-Ranking.
Composite31
Kontext256K
Input / 1M$2.5
Output / 1M$10
Wissensstichtag2025-06-01
Output-Speed61 t/s
TTFT (Default-API)0.36s
Statussolid
KnowledgeReasoningMathCodingAgenticLong contextTool use
Knowledge
71
Reasoning
53
Math
82
Coding
29
Agentic
15
Long context
18
Tool use
71

48average across 7 tested families

Benchmark-Platzierungen

Wo Command A auf jeder öffentlichen Benchmark-Quelle mit Zeile steht. Rang zählt jedes Modell mit neuester Zeile im selben Test — kein universeller Qualitätsscore.

BenchmarkFamilyRangScoreQuelleDatum
LiveCodeBenchCoding#12/ 1329Artificial Analysis2026-07-21
MATHMath#14/ 2082Artificial Analysis2026-07-21
AIMEMath#18/ 1913Artificial Analysis2026-07-21
MMLU-ProKnowledge#20/ 2171Artificial Analysis2026-07-21
τ³-Bench BankingAgentic#46/ 6415Artificial Analysis2026-07-21
AA output speedPerformance#47/ 69Artificial Analysis2026-07-21
AA time to first tokenPerformance#57/ 69Artificial Analysis2026-07-21
SciCodeCoding#61/ 6528Artificial Analysis2026-07-21
AA-LCRLong context#62/ 6418Artificial Analysis2026-07-21
Humanity's Last ExamReasoning#62/ 655Artificial Analysis2026-07-21
Artificial Analysis Intelligence IndexReasoning#65/ 678Artificial Analysis2026-07-21
GPQA DiamondReasoning#66/ 6953Artificial Analysis2026-07-21
Terminal-BenchAgentic#68/ 691Artificial Analysis2026-07-21
Family-Kontext

Die drei höchstscorierenden kartierten Modelle pro Capability-Family. Wo Command A auftaucht, ist es hervorgehoben.

KnowledgeCommand A · 71
  1. 1
    91
  2. 2
    91
  3. 3
    91
ReasoningCommand A · 53
  1. 1
    95
  2. 2
    94
  3. 3
    94
MathCommand A · 82
  1. 1
    99
  2. 2
    97
  3. 3
    96
CodingCommand A · 29
  1. 1
    96
  2. 2
    96
  3. 3
    96
AgenticCommand A · 15
  1. 1
    100
  2. 2
    100
  3. 3
    100
Long contextCommand A · 18
  1. 1
    75
  2. 2
    75
  3. 3
    74
Tool useCommand A · 71
  1. 1
    89
  2. 2
    87
  3. 3
    87
Neueste Belege

Benchmark-Zeilen im öffentlichen Ledger für Command A in den letzten 120 Tagen. Ältere Zeilen stehen in der vollen Tabelle unten.

14 neu
Family-Abdeckung

Ein hoher Composite, der eine schwache Family versteckt, ist eine Falle. Diese Balken zeigen Families ohne öffentlichen Test — und wo das Modell führt.

Knowledge71
Reasoning53
Math82
Coding29
Agentic15
Long context18
Tool use71
Preis vs. Performance
0255075100$0.10$0.30$1$3$10$30input USD / 1M tokens · log scale →← intelligence

Hover or tab any dot for name, score, and input price.

AnthropicOpenAIMistralAlibabaDeepSeekGoogleMoonshotMiniMaxxAIMetaZhipuXiaomiMeituanCohere

Volles Benchmark-Ledger

Jeder Katalog-Benchmark für Command A. Scores verlinken zur Originalquelle; Lücken heißen: noch keine öffentliche Zeile.

14 Quell-Zeilen

14 of 32 catalog benchmarks have a sourced row for Command A.

Knowledge

Reasoning

Math

Coding

Agentic

Long context

Tool use

Performance

Quellen