Model comparison

Granite 4.0 Micro vs Mistral Medium 3.1

Mistral Medium 3.1 is the stronger model overall, scoring 31.9 to 29.0 on the Noometry Index. Granite 4.0 Micro costs 20× less per token, which makes it the better buy when Mistral Medium 3.1's lead doesn't matter for your workload.

Last verified . 0 shared benchmarks.

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Mistral Medium 3.1 Mistral AI

31.9

Rank #266 Reported

Summary

  • The widest gap is in writing & preference, where Mistral Medium 3.1 leads 55.5 to 46.7.
  • Granite 4.0 Micro is cheaper at $0.017 / $0.11 per million input/output tokens, against $0.40 / $2 for Mistral Medium 3.1.
  • Mistral Medium 3.1 accepts more context: 131K tokens versus 131K.
  • Granite 4.0 Micro has downloadable open weights; the other is API-only.

Side by side

Granite 4.0 Micro and Mistral Medium 3.1 specifications
Granite 4.0 MicroMistral Medium 3.1
ProviderIBMMistral AI
Noometry Index29.031.9
Released2025-10-02—
WeightsOpenProprietary
Context window131K131K
Max output118K105K
Input $ / M tokens$0.017$0.40
Output $ / M tokens$0.11$2
Results tracked83

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Reasoning Granite 4.0 Micro leads

Granite 4.0 Micro: 19.2 (#265), Mistral Medium 3.1: 10.6 (#341)

Reasoning benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.1
NYT Connections (extended)—6.5%
Chess Puzzles0%—
Thematic Generalization—20.3%

Math Not comparable

Granite 4.0 Micro: 12.0 (#307), Mistral Medium 3.1: —

Math benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.1
OTIS Mock AIME 2024-20252.8%—
Omni-MATH20.9%—

Knowledge Not comparable

Granite 4.0 Micro: 9.9 (#304), Mistral Medium 3.1: —

Knowledge benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.1
GPQA Diamond28.3%—
MMLU-Pro39.5%—
GPQA (HELM)30.7%—

Instruction Following Not comparable

Granite 4.0 Micro: 69.9 (#169), Mistral Medium 3.1: —

Instruction Following benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.1
IFEval84.9%—

Writing & Preference Mistral Medium 3.1 leads

Granite 4.0 Micro: 46.7 (#216), Mistral Medium 3.1: 55.5 (#145)

Writing & Preference benchmarks
BenchmarkGranite 4.0 MicroMistral Medium 3.1
EQ-Bench Creative Writing—1476
WildBench67%—

Frequently asked questions

Is Granite 4.0 Micro better than Mistral Medium 3.1?

Mistral Medium 3.1 is the stronger model overall, scoring 31.9 to 29.0 on the Noometry Index. Granite 4.0 Micro costs 20× less per token, which makes it the better buy when Mistral Medium 3.1's lead doesn't matter for your workload.

Which is cheaper, Granite 4.0 Micro or Mistral Medium 3.1?

Granite 4.0 Micro is cheaper. It lists at $0.017 per million input tokens and $0.11 per million output tokens; Mistral Medium 3.1 lists at $0.40 and $2.

Which has the bigger context window?

Mistral Medium 3.1 does, with 131K tokens against 131K.

How many benchmarks do Granite 4.0 Micro and Mistral Medium 3.1 share?

0 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Mistral Medium 3.1 has 3.

Related comparisons

Go deeper