Model comparison

Granite 4.0 Micro vs Magistral Medium

Magistral Medium is the stronger model overall, scoring 35.2 to 29.0 on the Noometry Index. Granite 4.0 Micro costs 67× less per token, which makes it the better buy when Magistral Medium's lead doesn't matter for your workload.

Last verified . 0 shared benchmarks.

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

Summary

  • The widest gap is in knowledge, where Magistral Medium leads 33.5 to 9.9.
  • Granite 4.0 Micro is cheaper at $0.017 / $0.11 per million input/output tokens, against $2 / $5 for Magistral Medium.
  • Magistral Medium accepts more context: 262K tokens versus 131K.

Side by side

Granite 4.0 Micro and Magistral Medium specifications
Granite 4.0 MicroMagistral Medium
ProviderIBMMistral AI
Noometry Index29.035.2
Released2025-10-022025-03-17
WeightsOpenOpen
Context window131K262K
Max output118K16K
Input $ / M tokens$0.017$2
Output $ / M tokens$0.11$5
Results tracked822

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.0 Micro: —, Magistral Medium: 39.1 (#161)

Coding benchmarks
BenchmarkGranite 4.0 MicroMagistral Medium
SciCode—39.2%
LMArena Coding—1319

Reasoning Granite 4.0 Micro leads

Granite 4.0 Micro: 19.2 (#265), Magistral Medium: 8.6 (#348)

Reasoning benchmarks
BenchmarkGranite 4.0 MicroMagistral Medium
ARC-AGI-2—0%
Kagi LLM Benchmark—16.2%
ARC-AGI-1—6.1%
CritPt—0.3%
Chess Puzzles0%—
LMArena Hard Prompts—1267

Math Magistral Medium leads

Granite 4.0 Micro: 12.0 (#307), Magistral Medium: 35.1 (#189)

Math benchmarks
BenchmarkGranite 4.0 MicroMagistral Medium
OTIS Mock AIME 2024-20252.8%—
Omni-MATH20.9%—
LMArena Math—1250

Knowledge Magistral Medium leads

Granite 4.0 Micro: 9.9 (#304), Magistral Medium: 33.5 (#202)

Knowledge benchmarks
BenchmarkGranite 4.0 MicroMagistral Medium
GPQA Diamond28.3%—
MMLU-Pro39.5%—
GPQA (HELM)30.7%—
LMArena Expert—1223

Multilingual Not comparable

Granite 4.0 Micro: —, Magistral Medium: 39.6 (#224)

Multilingual benchmarks
BenchmarkGranite 4.0 MicroMagistral Medium
LMArena Non-English—1232
LMArena Chinese—1227
LMArena French—1267
LMArena German—1248
LMArena Japanese—1175
LMArena Korean—1125
LMArena Russian—1224
LMArena Spanish—1271

Instruction Following Granite 4.0 Micro leads

Granite 4.0 Micro: 69.9 (#169), Magistral Medium: 66.0 (#211)

Instruction Following benchmarks
BenchmarkGranite 4.0 MicroMagistral Medium
IFEval84.9%—
LMArena Instruction Following—1254

Long Context Not comparable

Granite 4.0 Micro: —, Magistral Medium: 39.3 (#183)

Long Context benchmarks
BenchmarkGranite 4.0 MicroMagistral Medium
LMArena Longer Query—1295

Writing & Preference Too close to call

Granite 4.0 Micro: 46.7 (#216), Magistral Medium: 46.3 (#219)

Writing & Preference benchmarks
BenchmarkGranite 4.0 MicroMagistral Medium
LMArena Text—1255
LMArena Creative Writing—1245
WildBench67%—
LMArena Multi-Turn—1275

Frequently asked questions

Is Granite 4.0 Micro better than Magistral Medium?

Magistral Medium is the stronger model overall, scoring 35.2 to 29.0 on the Noometry Index. Granite 4.0 Micro costs 67× less per token, which makes it the better buy when Magistral Medium's lead doesn't matter for your workload.

Which is cheaper, Granite 4.0 Micro or Magistral Medium?

Granite 4.0 Micro is cheaper. It lists at $0.017 per million input tokens and $0.11 per million output tokens; Magistral Medium lists at $2 and $5.

Which has the bigger context window?

Magistral Medium does, with 262K tokens against 131K.

How many benchmarks do Granite 4.0 Micro and Magistral Medium share?

0 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Magistral Medium has 22.

Related comparisons

Go deeper