Model comparison

Granite 4.0 Micro vs Mistral Medium 3.5

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 29.0 on the Noometry Index. Granite 4.0 Micro costs 74× less per token, which makes it the better buy when Mistral Medium 3.5's lead doesn't matter for your workload.

Last verified . 0 shared benchmarks.

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Mistral Medium 3.5 Mistral AI

40.2

Rank #152 Confirmed

Summary

  • The widest gap is in knowledge, where Mistral Medium 3.5 leads 40.0 to 9.9.
  • Granite 4.0 Micro is cheaper at $0.017 / $0.11 per million input/output tokens, against $1.50 / $7.50 for Mistral Medium 3.5.
  • Mistral Medium 3.5 accepts more context: 262K tokens versus 131K.

Side by side

Granite 4.0 Micro and Mistral Medium 3.5 specifications
Granite 4.0 MicroMistral Medium 3.5
ProviderIBMMistral AI
Noometry Index29.040.2
Released2025-10-02—
WeightsOpenOpen
Context window131K262K
Max output118K210K
Input $ / M tokens$0.017$1.50
Output $ / M tokens$0.11$7.50
Results tracked822

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.0 Micro: —, Mistral Medium 3.5: 36.0 (#213)

Coding benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.5
LMArena WebDev—1264
LMArena Coding—1461

Reasoning Granite 4.0 Micro leads

Granite 4.0 Micro: 19.2 (#265), Mistral Medium 3.5: 17.3 (#295)

Reasoning benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.5
Kagi LLM Benchmark—41.4%
NYT Connections (extended)—12.9%
Chess Puzzles0%—
LMArena Hard Prompts—1436
Epoch Capabilities Index—141.35

Math Mistral Medium 3.5 leads

Granite 4.0 Micro: 12.0 (#307), Mistral Medium 3.5: 39.1 (#113)

Math benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.5
OTIS Mock AIME 2024-20252.8%—
Omni-MATH20.9%—
LMArena Math—1431

Knowledge Mistral Medium 3.5 leads

Granite 4.0 Micro: 9.9 (#304), Mistral Medium 3.5: 40.0 (#126)

Knowledge benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.5
GPQA Diamond28.3%—
MMLU-Pro39.5%—
GPQA (HELM)30.7%—
LMArena Expert—1432

Multimodal Not comparable

Granite 4.0 Micro: —, Mistral Medium 3.5: 38.3 (#65)

Multimodal benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.5
LMArena Vision—1223

Multilingual Not comparable

Granite 4.0 Micro: —, Mistral Medium 3.5: 51.9 (#100)

Multilingual benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.5
LMArena Non-English—1404
LMArena Chinese—1442
LMArena French—1448
LMArena German—1451
LMArena Korean—1385
LMArena Russian—1395
LMArena Spanish—1409

Instruction Following Mistral Medium 3.5 leads

Granite 4.0 Micro: 69.9 (#169), Mistral Medium 3.5: 74.6 (#90)

Instruction Following benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.5
IFEval84.9%—
LMArena Instruction Following—1415

Long Context Not comparable

Granite 4.0 Micro: —, Mistral Medium 3.5: 43.2 (#103)

Long Context benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.5
LMArena Longer Query—1415

Writing & Preference Mistral Medium 3.5 leads

Granite 4.0 Micro: 46.7 (#216), Mistral Medium 3.5: 58.5 (#117)

Writing & Preference benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.5
LMArena Text—1421
LMArena Creative Writing—1374
WildBench67%—
EQ-Bench 4—993
LMArena Multi-Turn—1423

Frequently asked questions

Is Granite 4.0 Micro better than Mistral Medium 3.5?

Mistral Medium 3.5 is the stronger model overall, scoring 40.2 to 29.0 on the Noometry Index. Granite 4.0 Micro costs 74× less per token, which makes it the better buy when Mistral Medium 3.5's lead doesn't matter for your workload.

Which is cheaper, Granite 4.0 Micro or Mistral Medium 3.5?

Granite 4.0 Micro is cheaper. It lists at $0.017 per million input tokens and $0.11 per million output tokens; Mistral Medium 3.5 lists at $1.50 and $7.50.

Which has the bigger context window?

Mistral Medium 3.5 does, with 262K tokens against 131K.

How many benchmarks do Granite 4.0 Micro and Mistral Medium 3.5 share?

0 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Mistral Medium 3.5 has 22.

Related comparisons

Go deeper