Pick two models. See the diff.
Mistral Large 3 vs Llama 4 Maverick
Spec showdown
Capacity, price, freshness.
Composite scores are within 5 points — 48 for Mistral Large 3 vs 49 for Llama 4 Maverick. The choice comes down to capability profile and price. Mistral Large 3 has the higher Vals/AA intelligence score (16 vs 14). Llama 4 Maverick is more affordable at $0.17/1M input — about 3× cheaper than $0.50/1M. Llama 4 Maverick's 1M context window dwarfs Mistral Large 3's 262K — a meaningful difference for long-document or codebase work.
- Composite scoreMistral Large 348Llama 4 Maverick49Winner on this row
- Intelligence (Vals/AA)Mistral Large 316Winner on this rowLlama 4 Maverick14
- Context windowMistral Large 3262KLlama 4 Maverick1MWinner on this row
- Max outputMistral Large 38KLlama 4 Maverick16KWinner on this row
- Input price / 1M (lower wins)Mistral Large 3$0.50Llama 4 Maverick$0.17Winner on this row
- Output speed (AA default API)Mistral Large 364 t/sLlama 4 Maverick110 t/sWinner on this row
- Time to first token (AA default API)Mistral Large 30.59sWinner on this rowLlama 4 Maverick0.60s
Capability grid
What each model can do.
Modalities
Access
Family radar
Top score per benchmark family.
Benchmark ranks
Every shared benchmark, side by side.
Knowledge
Reasoning
Math
Coding
Agentic
Long context
Tool use
Performance
Dossier facts
Dates, sources, fineprint.
Ecosystem
Which tools wrap each model.
Llama 4 Maverick
Not listed in any tool yet.
Verdict
Which model to actually pick.
Composite scores are within 5 points — 48 for Mistral Large 3 vs 49 for Llama 4 Maverick. The choice comes down to capability profile and price. Mistral Large 3 has the higher Vals/AA intelligence score (16 vs 14). Llama 4 Maverick is more affordable at $0.17/1M input — about 3× cheaper than $0.50/1M. Llama 4 Maverick's 1M context window dwarfs Mistral Large 3's 262K — a meaningful difference for long-document or codebase work.