Model comparison

Claude 2 vs Magistral Medium

Magistral Medium is the stronger model overall, scoring 35.2 to 25.0 on the Noometry Index.

Last verified . 0 shared benchmarks.

Claude 2 Anthropic

25.0

Rank #346 Reported

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

Summary

  • The widest gap is in math, where Magistral Medium leads 35.1 to 9.3.
  • Magistral Medium has downloadable open weights; the other is API-only.

Side by side

Claude 2 and Magistral Medium specifications
Claude 2Magistral Medium
ProviderAnthropicMistral AI
Noometry Index25.035.2
Released2023-07-112025-03-17
WeightsProprietaryOpen
Context window—262K
Max output—16K
Input $ / M tokens—$2
Output $ / M tokens—$5
Results tracked822

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Claude 2: —, Magistral Medium: 39.1 (#161)

Coding benchmarks
BenchmarkClaude 2Magistral Medium
SciCode—39.2%
LMArena Coding—1319
HumanEval+61.6%—

Reasoning Claude 2 leads

Claude 2: 21.7 (#216), Magistral Medium: 8.6 (#348)

Reasoning benchmarks
BenchmarkClaude 2Magistral Medium
ARC-AGI-2—0%
Kagi LLM Benchmark—16.2%
ARC-AGI-1—6.1%
CritPt—0.3%
LMArena Hard Prompts—1267
DTBench51.9%—
Epoch Capabilities Index120.13—

Math Magistral Medium leads

Claude 2: 9.3 (#320), Magistral Medium: 35.1 (#189)

Math benchmarks
BenchmarkClaude 2Magistral Medium
OTIS Mock AIME 2024-20252.5%—
LMArena Math—1250
MATH Level 511.7%—

Knowledge Magistral Medium leads

Claude 2: 16.9 (#287), Magistral Medium: 33.5 (#202)

Knowledge benchmarks
BenchmarkClaude 2Magistral Medium
GPQA Diamond34.7%—
LMArena Expert—1223
MMLU78.5%—
TriviaQA87.5%—

Multilingual Not comparable

Claude 2: —, Magistral Medium: 39.6 (#224)

Multilingual benchmarks
BenchmarkClaude 2Magistral Medium
LMArena Non-English—1232
LMArena Chinese—1227
LMArena French—1267
LMArena German—1248
LMArena Japanese—1175
LMArena Korean—1125
LMArena Russian—1224
LMArena Spanish—1271

Instruction Following Not comparable

Claude 2: —, Magistral Medium: 66.0 (#211)

Instruction Following benchmarks
BenchmarkClaude 2Magistral Medium
LMArena Instruction Following—1254

Long Context Not comparable

Claude 2: —, Magistral Medium: 39.3 (#183)

Long Context benchmarks
BenchmarkClaude 2Magistral Medium
LMArena Longer Query—1295

Writing & Preference Not comparable

Claude 2: —, Magistral Medium: 46.3 (#219)

Writing & Preference benchmarks
BenchmarkClaude 2Magistral Medium
LMArena Text—1255
LMArena Creative Writing—1245
LMArena Multi-Turn—1275

Frequently asked questions

Is Claude 2 better than Magistral Medium?

Magistral Medium is the stronger model overall, scoring 35.2 to 25.0 on the Noometry Index.

How many benchmarks do Claude 2 and Magistral Medium share?

0 benchmarks have published results for both models. Claude 2 has 8 scored results on Noometry and Magistral Medium has 22.

Related comparisons

Go deeper