Model comparison

Codellama 70b Instruct vs Magistral Medium

Magistral Medium is the stronger model overall, scoring 35.2 to 33.7 on the Noometry Index.

Last verified . 4 shared benchmarks.

Codellama 70b Instruct Meta

33.7

Rank #237 Confirmed

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

Summary

  • They share 4 benchmarks with published results for both. Codellama 70b Instruct scores higher in 1 category and Magistral Medium in 4 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in multilingual, where Magistral Medium leads 39.6 to 24.8.

Side by side

Codellama 70b Instruct and Magistral Medium specifications
Codellama 70b InstructMagistral Medium
ProviderMetaMistral AI
Noometry Index33.735.2
Released—2025-03-17
WeightsOpenOpen
Context window—262K
Max output—16K
Input $ / M tokens—$2
Output $ / M tokens—$5
Results tracked722

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Medium leads

Codellama 70b Instruct: 37.6 (#193), Magistral Medium: 39.1 (#161)

Coding benchmarks
BenchmarkCodellama 70b InstructMagistral Medium
SciCode—39.2%
BigCodeBench Instruct40.7%—
LMArena Coding—1319
BigCodeBench Complete49.6%—
HumanEval+65.9%—

Reasoning Codellama 70b Instruct leads

Codellama 70b Instruct: 20.1 (#242), Magistral Medium: 8.6 (#348)

Reasoning benchmarks
BenchmarkCodellama 70b InstructMagistral Medium
LMArena Hard Prompts10521267
ARC-AGI-2—0%
Kagi LLM Benchmark—16.2%
ARC-AGI-1—6.1%
CritPt—0.3%

Math Not comparable

Codellama 70b Instruct: —, Magistral Medium: 35.1 (#189)

Math benchmarks
BenchmarkCodellama 70b InstructMagistral Medium
LMArena Math—1250

Knowledge Not comparable

Codellama 70b Instruct: —, Magistral Medium: 33.5 (#202)

Knowledge benchmarks
BenchmarkCodellama 70b InstructMagistral Medium
LMArena Expert—1223

Multilingual Magistral Medium leads

Codellama 70b Instruct: 24.8 (#288), Magistral Medium: 39.6 (#224)

Multilingual benchmarks
BenchmarkCodellama 70b InstructMagistral Medium
LMArena Non-English9921232
LMArena Chinese—1227
LMArena French—1267
LMArena German—1248
LMArena Japanese—1175
LMArena Korean—1125
LMArena Russian—1224
LMArena Spanish—1271

Instruction Following Magistral Medium leads

Codellama 70b Instruct: 51.9 (#293), Magistral Medium: 66.0 (#211)

Instruction Following benchmarks
BenchmarkCodellama 70b InstructMagistral Medium
LMArena Instruction Following10241254

Long Context Not comparable

Codellama 70b Instruct: —, Magistral Medium: 39.3 (#183)

Long Context benchmarks
BenchmarkCodellama 70b InstructMagistral Medium
LMArena Longer Query—1295

Writing & Preference Magistral Medium leads

Codellama 70b Instruct: 33.4 (#277), Magistral Medium: 46.3 (#219)

Writing & Preference benchmarks
BenchmarkCodellama 70b InstructMagistral Medium
LMArena Text10571255
LMArena Creative Writing—1245
LMArena Multi-Turn—1275

Frequently asked questions

Is Codellama 70b Instruct better than Magistral Medium?

Magistral Medium is the stronger model overall, scoring 35.2 to 33.7 on the Noometry Index.

Is Codellama 70b Instruct or Magistral Medium better for coding?

Magistral Medium scores higher on coding benchmarks: 39.1 versus 37.6 in the Noometry coding category.

How many benchmarks do Codellama 70b Instruct and Magistral Medium share?

4 benchmarks have published results for both models. Codellama 70b Instruct has 7 scored results on Noometry and Magistral Medium has 22.

Related comparisons

Go deeper