Model comparison

Granite 4.0 Micro vs Mistral 7B

Granite 4.0 Micro is the stronger model overall, scoring 29.0 to 23.0 on the Noometry Index.

Last verified . 3 shared benchmarks.

Granite 4.0 Micro IBM

29.0

Rank #318 Confirmed

Mistral 7B Mistral AI

23.0

Rank #351 Confirmed

Summary

  • They share 3 benchmarks with published results for both. Granite 4.0 Micro scores higher in 5 categories and Mistral 7B in 0 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Granite 4.0 Micro leads 46.7 to 30.7.
  • The biggest single-benchmark swing is GPQA Diamond: 28.3% for Granite 4.0 Micro and 15.2% for Mistral 7B.
  • Granite 4.0 Micro is cheaper at $0.017 / $0.11 per million input/output tokens, against $0.25 / $0.25 for Mistral 7B.
  • Granite 4.0 Micro accepts more context: 131K tokens versus 8K.

Side by side

Granite 4.0 Micro and Mistral 7B specifications
Granite 4.0 MicroMistral 7B
ProviderIBMMistral AI
Noometry Index29.023.0
Released2025-10-022023-09-27
WeightsOpenOpen
Context window131K8K
Max output118K8K
Input $ / M tokens$0.017$0.25
Output $ / M tokens$0.11$0.25
Results tracked837

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Granite 4.0 Micro: —, Mistral 7B: 26.4 (#326)

Coding benchmarks
BenchmarkGranite 4.0 MicroMistral 7B
BigCodeBench Instruct—19.5%
LMArena Coding—1082
BigCodeBench Complete—27.3%
HumanEval+—36%
MBPP+—42.1%

Reasoning Granite 4.0 Micro leads

Granite 4.0 Micro: 19.2 (#265), Mistral 7B: 13.1 (#336)

Reasoning benchmarks
BenchmarkGranite 4.0 MicroMistral 7B
Chess Puzzles0%0%
LMArena Hard Prompts—1067
DTBench—42.5%
Adversarial NLI—47.1%
BIG-Bench Hard—56.1%
Epoch Capabilities Index—112.21
HellaSwag—81%
PIQA—83%
WinoGrande—75.3%

Math Granite 4.0 Micro leads

Granite 4.0 Micro: 12.0 (#307), Mistral 7B: 8.1 (#325)

Math benchmarks
BenchmarkGranite 4.0 MicroMistral 7B
OTIS Mock AIME 2024-20252.8%0.3%
Omni-MATH20.9%—
LMArena Math—1085
MATH Level 5—3.7%
GSM8K—54.4%

Knowledge Granite 4.0 Micro leads

Granite 4.0 Micro: 9.9 (#304), Mistral 7B: 7.4 (#311)

Knowledge benchmarks
BenchmarkGranite 4.0 MicroMistral 7B
GPQA Diamond28.3%15.2%
MMLU-Pro39.5%—
GPQA (HELM)30.7%—
LMArena Expert—1036
ARC (AI2) Challenge—78.6%
BoolQ—87.4%
MMLU—62.5%
OpenBookQA—79.8%
TriviaQA—75.2%

Multilingual Not comparable

Granite 4.0 Micro: —, Mistral 7B: 25.8 (#283)

Multilingual benchmarks
BenchmarkGranite 4.0 MicroMistral 7B
LMArena Non-English—1012
LMArena Chinese—1009
LMArena French—1037
LMArena German—987
LMArena Japanese—878
LMArena Russian—1018
LMArena Spanish—1026

Instruction Following Granite 4.0 Micro leads

Granite 4.0 Micro: 69.9 (#169), Mistral 7B: 54.2 (#280)

Instruction Following benchmarks
BenchmarkGranite 4.0 MicroMistral 7B
IFEval84.9%—
LMArena Instruction Following—1060

Long Context Not comparable

Granite 4.0 Micro: —, Mistral 7B: 32.2 (#271)

Long Context benchmarks
BenchmarkGranite 4.0 MicroMistral 7B
LMArena Longer Query—1060

Writing & Preference Granite 4.0 Micro leads

Granite 4.0 Micro: 46.7 (#216), Mistral 7B: 30.7 (#286)

Writing & Preference benchmarks
BenchmarkGranite 4.0 MicroMistral 7B
LMArena Text—1090
LMArena Creative Writing—1068
WildBench67%—
LMArena Multi-Turn—1062

Frequently asked questions

Is Granite 4.0 Micro better than Mistral 7B?

Granite 4.0 Micro is the stronger model overall, scoring 29.0 to 23.0 on the Noometry Index.

Which is cheaper, Granite 4.0 Micro or Mistral 7B?

Granite 4.0 Micro is cheaper. It lists at $0.017 per million input tokens and $0.11 per million output tokens; Mistral 7B lists at $0.25 and $0.25.

Which has the bigger context window?

Granite 4.0 Micro does, with 131K tokens against 8K.

How many benchmarks do Granite 4.0 Micro and Mistral 7B share?

3 benchmarks have published results for both models. Granite 4.0 Micro has 8 scored results on Noometry and Mistral 7B has 37.

Related comparisons

Go deeper