Model comparison

Claude Haiku 4.5 vs Mistral Medium

Claude Haiku 4.5 is the stronger model overall, scoring 39.5 to 36.3 on the Noometry Index.

Last verified . 29 shared benchmarks.

Claude Haiku 4.5 Anthropic

39.5

Rank #165 Confirmed

Mistral Medium Mistral AI

36.3

Rank #218 Confirmed

Summary

  • They share 29 benchmarks with published results for both. Claude Haiku 4.5 scores higher in 5 categories and Mistral Medium in 5 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in math, where Claude Haiku 4.5 leads 44.9 to 28.1.
  • The biggest single-benchmark swing is OTIS Mock AIME 2024-2025: 66.7% for Claude Haiku 4.5 and 32.2% for Mistral Medium.
  • Claude Haiku 4.5 is cheaper at $1 / $5 per million input/output tokens, against $1.50 / $7.50 for Mistral Medium.
  • Mistral Medium accepts more context: 262K tokens versus 200K.
  • Mistral Medium has downloadable open weights; the other is API-only.

Side by side

Claude Haiku 4.5 and Mistral Medium specifications
Claude Haiku 4.5Mistral Medium
ProviderAnthropicMistral AI
Noometry Index39.536.3
Released2025-10-152023-12-11
WeightsProprietaryOpen
Context window200K262K
Max output64K262K
Input $ / M tokens$1$1.50
Output $ / M tokens$5$7.50
Results tracked5336

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Claude Haiku 4.5 leads

Claude Haiku 4.5: 44.0 (#78), Mistral Medium: 34.2 (#243)

Coding benchmarks
BenchmarkClaude Haiku 4.5Mistral Medium
SciCode43.3%40.2%
WeirdML45.4%43.7%
LMArena Coding14531434
ALE-Bench653.48763.98
FrontierCode—8%
SWE-bench Verified (bash only)66.6%—
LMArena WebDev1330—
SWE-bench Multilingual64.7%—

Agentic & Tool Use Claude Haiku 4.5 leads

Claude Haiku 4.5: 33.6 (#52), Mistral Medium: 28.3 (#90)

Agentic & Tool Use benchmarks
BenchmarkClaude Haiku 4.5Mistral Medium
Berkeley Function Calling Leaderboard68.7%37.7%
Terminal-Bench35.5%—
DeepResearch Bench45.5%—
BALROG31.2%—
ExploitBench13.7%—
Vending-Bench 2458.89—

Reasoning Mistral Medium leads

Claude Haiku 4.5: 15.1 (#320), Mistral Medium: 24.0 (#167)

Reasoning benchmarks
BenchmarkClaude Haiku 4.5Mistral Medium
CritPt0%0%
LMArena Hard Prompts14201426
DTBench73.6%75.5%
LMCA30.9%26.1%
ARC-AGI-24%—
Kagi LLM Benchmark—50%
NYT Connections (extended)14.3%—
ARC-AGI-147.7%—
Chess Puzzles8%—
Surface Evolver Bench—26.9%
Epoch Capabilities Index142.41—
ForecastBench61.4—

Math Claude Haiku 4.5 leads

Claude Haiku 4.5: 44.9 (#78), Mistral Medium: 28.1 (#245)

Math benchmarks
BenchmarkClaude Haiku 4.5Mistral Medium
OTIS Mock AIME 2024-202566.7%32.2%
LMArena Math13961408
MATH Level 596.4%81.6%
FrontierMath (Feb 2025 set)5.9%0.3%
ProofBench—9%
Omni-MATH56.1%—
FrontierMath Tier 4 (v1)2.1%—

Knowledge Claude Haiku 4.5 leads

Claude Haiku 4.5: 37.7 (#153), Mistral Medium: 25.0 (#265)

Knowledge benchmarks
BenchmarkClaude Haiku 4.5Mistral Medium
GPQA Diamond71.2%59.5%
Vectara Hallucination Rate9.8%22.7%
LMArena Expert14421408
Humanity's Last Exam—4.5%
SimpleQA Verified13.2%—
MMLU-Pro77.7%—
GPQA (HELM)60.5%—

Multimodal Mistral Medium leads

Claude Haiku 4.5: 26.8 (#118), Mistral Medium: 35.3 (#88)

Multimodal benchmarks
BenchmarkClaude Haiku 4.5Mistral Medium
LMArena Vision—1172
Blueprint-Bench 20%—
LMArena Document1420—

Multilingual Mistral Medium leads

Claude Haiku 4.5: 49.9 (#129), Mistral Medium: 52.1 (#91)

Multilingual benchmarks
BenchmarkClaude Haiku 4.5Mistral Medium
LMArena Non-English13771408
LMArena Chinese14171447
LMArena French14081459
LMArena German13751432
LMArena Japanese13391378
LMArena Korean13471380
LMArena Russian13811411
LMArena Spanish14201433

Instruction Following Mistral Medium leads

Claude Haiku 4.5: 71.4 (#149), Mistral Medium: 73.7 (#116)

Instruction Following benchmarks
BenchmarkClaude Haiku 4.5Mistral Medium
LMArena Instruction Following14141398
IFEval80.1%—

Long Context Too close to call

Claude Haiku 4.5: 43.6 (#92), Mistral Medium: 42.9 (#114)

Long Context benchmarks
BenchmarkClaude Haiku 4.5Mistral Medium
LMArena Longer Query14271406

Writing & Preference Mistral Medium leads

Claude Haiku 4.5: 57.9 (#123), Mistral Medium: 60.0 (#103)

Writing & Preference benchmarks
BenchmarkClaude Haiku 4.5Mistral Medium
LMArena Text13961424
LMArena Creative Writing13721391
LMArena Multi-Turn14091418
Short-Story Creative Writing—77.3%
WildBench83.9%—
EQ-Bench 41064—

Frequently asked questions

Is Claude Haiku 4.5 better than Mistral Medium?

Claude Haiku 4.5 is the stronger model overall, scoring 39.5 to 36.3 on the Noometry Index.

Which is cheaper, Claude Haiku 4.5 or Mistral Medium?

Claude Haiku 4.5 is cheaper. It lists at $1 per million input tokens and $5 per million output tokens; Mistral Medium lists at $1.50 and $7.50.

Is Claude Haiku 4.5 or Mistral Medium better for coding?

Claude Haiku 4.5 scores higher on coding benchmarks: 44.0 versus 34.2 in the Noometry coding category.

Which has the bigger context window?

Mistral Medium does, with 262K tokens against 200K.

How many benchmarks do Claude Haiku 4.5 and Mistral Medium share?

29 benchmarks have published results for both models. Claude Haiku 4.5 has 53 scored results on Noometry and Mistral Medium has 36.

Related comparisons

Go deeper