Model comparison

Granite 4.0 Micro vs Mistral Large 4

Mistral Large 4 is the stronger model overall, scoring 43.1 to 29.0 on the Noometry Index. Granite 4.0 Micro costs 25× less per token, which makes it the better buy when Mistral Large 4's lead doesn't matter for your workload.

Last verified . 0 shared benchmarks.

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Mistral Large 4 Mistral AI

43.1

Rank #99 Confirmed

Summary

  • The widest gap is in math, where Mistral Large 4 leads 40.4 to 12.0.
  • Granite 4.0 Micro is cheaper at $0.017 / $0.11 per million input/output tokens, against $0.68 / $2.09 for Mistral Large 4.
  • Mistral Large 4 accepts more context: 1.05M tokens versus 131K.
  • Granite 4.0 Micro has downloadable open weights; the other is API-only.

Side by side

Granite 4.0 Micro and Mistral Large 4 specifications
Granite 4.0 MicroMistral Large 4
ProviderIBMMistral AI
Noometry Index29.043.1
Released2025-10-022026-10-06
WeightsOpenProprietary
Context window131K1.05M
Max output118K262K
Input $ / M tokens$0.017$0.68
Output $ / M tokens$0.11$2.09
Results tracked815

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.0 Micro: —, Mistral Large 4: 48.6 (#57)

Coding benchmarks
BenchmarkGranite 4.0 MicroMistral Large 4
LMArena WebDev—1541
LMArena Coding—1475

Reasoning Mistral Large 4 leads

Granite 4.0 Micro: 19.2 (#265), Mistral Large 4: 22.5 (#192)

Reasoning benchmarks
BenchmarkGranite 4.0 MicroMistral Large 4
NYT Connections (extended)—27.4%
Chess Puzzles0%—
LMArena Hard Prompts—1444

Math Mistral Large 4 leads

Granite 4.0 Micro: 12.0 (#307), Mistral Large 4: 40.4 (#91)

Math benchmarks
BenchmarkGranite 4.0 MicroMistral Large 4
OTIS Mock AIME 2024-20252.8%—
Omni-MATH20.9%—
LMArena Math—1488

Knowledge Mistral Large 4 leads

Granite 4.0 Micro: 9.9 (#304), Mistral Large 4: 36.6 (#166)

Knowledge benchmarks
BenchmarkGranite 4.0 MicroMistral Large 4
GPQA Diamond28.3%—
SimpleQA Verified—20%
MMLU-Pro39.5%—
GPQA (HELM)30.7%—
LMArena Expert—1447

Multilingual Not comparable

Granite 4.0 Micro: —, Mistral Large 4: 52.6 (#82)

Multilingual benchmarks
BenchmarkGranite 4.0 MicroMistral Large 4
LMArena Non-English—1415
LMArena Chinese—1491
LMArena Russian—1414

Instruction Following Mistral Large 4 leads

Granite 4.0 Micro: 69.9 (#169), Mistral Large 4: 75.0 (#76)

Instruction Following benchmarks
BenchmarkGranite 4.0 MicroMistral Large 4
IFEval84.9%—
LMArena Instruction Following—1424

Long Context Not comparable

Granite 4.0 Micro: —, Mistral Large 4: 43.6 (#89)

Long Context benchmarks
BenchmarkGranite 4.0 MicroMistral Large 4
LMArena Longer Query—1429

Writing & Preference Mistral Large 4 leads

Granite 4.0 Micro: 46.7 (#216), Mistral Large 4: 60.4 (#97)

Writing & Preference benchmarks
BenchmarkGranite 4.0 MicroMistral Large 4
LMArena Text—1427
LMArena Creative Writing—1361
WildBench67%—
LMArena Multi-Turn—1424

Frequently asked questions

Is Granite 4.0 Micro better than Mistral Large 4?

Mistral Large 4 is the stronger model overall, scoring 43.1 to 29.0 on the Noometry Index. Granite 4.0 Micro costs 25× less per token, which makes it the better buy when Mistral Large 4's lead doesn't matter for your workload.

Which is cheaper, Granite 4.0 Micro or Mistral Large 4?

Granite 4.0 Micro is cheaper. It lists at $0.017 per million input tokens and $0.11 per million output tokens; Mistral Large 4 lists at $0.68 and $2.09.

Which has the bigger context window?

Mistral Large 4 does, with 1.05M tokens against 131K.

How many benchmarks do Granite 4.0 Micro and Mistral Large 4 share?

0 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Mistral Large 4 has 15.

Related comparisons

Go deeper