Models / Mistral

Mistral Medium 3.5

Mistral Medium · Released 2026-04-29

AA Intelligence Index 29.9. Listed on Artificial Analysis with API pricing $1.5/$7.5 per 1M tokens. Benchmark rows sync from the AA snapshot — see the ledger on this page.

Frontier model tracked on Artificial Analysis. Use the benchmark ledger below for sourced scores; we do not infer numbers beyond published rows.

#97 of 131 on AA Index · current snapshot
30 AA Index · 10 benchmark rows · 1 sourcesPosition is benchmark-specific — not a cross-family or cross-source ranking.
Composite44
ContextUnlisted
Input / 1M$1.5
Output / 1M$7.5
Knowledge cutoffunpublished
Output speed101 t/s
TTFT (default API)0.59s
Statussolid
ReasoningCodingAgenticLong context
Reasoning
75
Coding
40
Agentic
51
Long context
61

57average across 4 tested families

Benchmark placements

Where Mistral Medium 3.5 places on each public benchmark source that publishes a row. Rank counts every model with a latest row in the same test — not a universal quality score.

BenchmarkFamilyRankScoreSourceDate
AA output speedPerformance#38/ 149Artificial Analysis2026-07-28
IFBenchReasoning#58/ 12169Artificial Analysis2026-07-28
AA time to first tokenPerformance#68/ 149Artificial Analysis2026-07-28
Terminal-BenchAgentic#72/ 14851Artificial Analysis2026-07-28
AA-LCRLong context#94/ 14361Artificial Analysis2026-07-28
SciCodeCoding#104/ 14440Artificial Analysis2026-07-28
τ³-Bench BankingAgentic#105/ 14314Artificial Analysis2026-07-28
Artificial Analysis Intelligence IndexReasoning#111/ 14730Artificial Analysis2026-07-28
Humanity's Last ExamReasoning#114/ 14413Artificial Analysis2026-07-28
GPQA DiamondReasoning#125/ 14975Artificial Analysis2026-07-28
Family context

The three highest-scoring carded models in each capability family. Where Mistral Medium 3.5 shows up, it's highlighted.

ReasoningMistral Medium 3.5 · 75
  1. 1
    95
  2. 2
    94
  3. 3
    94
CodingMistral Medium 3.5 · 40
  1. 1
    96
  2. 2
    96
  3. 3
    96
AgenticMistral Medium 3.5 · 51
  1. 1
    100
  2. 2
    100
  3. 3
    100
Long contextMistral Medium 3.5 · 61
  1. 1
    76
  2. 2
    76
  3. 3
    75
Newest receipts

Benchmark rows added to the public ledger for Mistral Medium 3.5 in the last 120 days. Older rows live in the full table below.

10 recent
Family coverage

A high composite that hides a weak family is a trap. These bars surface the families where this model hasn't been publicly tested, and where it leads.

Reasoning75
Coding40
Agentic51
Long context61
Price vs. performance
0255075100$0.10$0.30$1$3$10$30input USD / 1M tokens · log scale →← intelligence

Hover or tab any dot for name, score, and input price.

AnthropicOpenAIMistralAlibabaDeepSeekGoogleMoonshotxAIMiniMaxMetaZhipuXiaomiOtherMeituanCohereNVIDIA

Full benchmark ledger

Every catalog benchmark for Mistral Medium 3.5. Scores link to the original source; gaps mean no public row exists yet.

10 sourced rows

10 of 32 catalog benchmarks have a sourced row for Mistral Medium 3.5.

Reasoning

Coding

Agentic

Long context

Performance

Sources