Mistral AI, open weights

Mistral Small 3

Mistral Small 3 by Mistral AI ranks 278th of 354 ranked models on the Noometry Index as of October 2026, with a score of 31.2. Its strongest category is coding, where it ranks 207th. API pricing starts at $0.05 per million input tokens and $0.08 per million output tokens, with a 33K-token context window.

Last verified

Specifications

Noometry rank
#278 of 354
Index score
31.2
Evidence
Confirmed 24 results
Provider
Mistral AI
Released
January 30, 2025
Weights
Open weights
Reasoning
Unknown
Context window
33K
Max output
16K
Input price
$0.05 / M
Output price
$0.08 / M
Blended price
$0.0575 / M
Output speed
Not measured
Value
#5 of 219
Knowledge cutoff
Unknown
Input
text

Category scores

Each category score combines every public result we have in that category.

Mistral Small 3 category scores
  1. Coding 36.5
  2. Reasoning 18.9
  3. Math 16.3
  4. Knowledge 25.1
  5. Multilingual 37.3
  6. Instruction Following 63.7
  7. Long Context 37.8
  8. Writing & Preference 32.2
Mistral Small 3 category ranks
CategoryScoreRankResults
Coding36.5#2073
Reasoning18.9#2732
Math16.3#2952
Knowledge25.1#2633
Multilingual37.3#2361
Instruction Following63.7#2291
Long Context37.8#2111
Writing & Preference32.2#2804

Strengths and weaknesses

Categories where Mistral Small 3 places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Mistral Small 3: strongest categories
CategoryScorevs medianRank
Coding36.5−2.2#207 of 340, top 61%
Long Context37.8−3.1#211 of 296, top 72%
Instruction Following63.7−7.6#229 of 305, top 76%

Weakest categories

Mistral Small 3: weakest categories
CategoryScorevs medianRank
Math16.3−20.3#295 of 327, top 91%
Writing & Preference32.2−21.6#280 of 312, top 90%
Knowledge25.1−12.2#263 of 314, top 84%

Closest competitors

The models ranked just above and below Mistral Small 3. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Mistral Small 3
ModelRankScoreBlended $/MSpeed
Wizardlm 13b#27431.4——Compare
Qwen-14B#27531.4——Compare
Phi 3 Mini 4k Instruct June 2024#27631.3——Compare
Gemma 1.1 7b IT#27731.3——Compare
Phi-4#27931.2$0.0875—Compare
Mistral Small 3.2#28031.2$0.1368Compare
Amazon Nova Pro#28131.0$1.40—Compare
Llama 4 Maverick#28230.9$0.30456Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Mistral Small 3 Coding benchmark results
BenchmarkScorePositionSettingSourceDate
BigCodeBench Instruct45.3%#20 of 64, top 32%BigCodeBench2025-01-31
LMArena Coding1246#227 of 294, top 78%LMArena2026-10-08
BigCodeBench Complete50.4%#32 of 66, top 49%BigCodeBench2025-01-31

Reasoning

Mistral Small 3 Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Chess Puzzles0%#124 of 129, top 97%Epoch AI2026-08-30
LMArena Hard Prompts1233#228 of 297, top 77%LMArena2026-10-08
Epoch Capabilities Index127.07#151 of 213, top 71%Epoch AI2025-01-30

Math

Mistral Small 3 Math benchmark results
BenchmarkScorePositionSettingSourceDate
OTIS Mock AIME 2024-20256.7%#143 of 173, top 83%Epoch AI2026-08-30
LMArena Math1240#220 of 285, top 78%LMArena2026-10-08

Knowledge

Mistral Small 3 Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond47.3%#137 of 186, top 74%Epoch AI2026-08-30
Confabulations (lower is better)25.2%#42 of 51, top 83%Lech Mazur benchmarks
LMArena Expert1202#221 of 273, top 81%LMArena2026-10-08

Multilingual

Mistral Small 3 Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1198#236 of 297, top 80%LMArena2026-10-08
LMArena Chinese1204#232 of 285, top 82%LMArena2026-10-08
LMArena French1203#192 of 223, top 87%LMArena2026-10-08
LMArena German1211#188 of 231, top 82%LMArena2026-10-08
LMArena Japanese1111#185 of 211, top 88%LMArena2026-10-08
LMArena Korean1188#166 of 213, top 78%LMArena2026-10-08
LMArena Russian1216#229 of 283, top 81%LMArena2026-10-08

Instruction Following

Mistral Small 3 Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1214#230 of 298, top 78%LMArena2026-10-08

Long Context

Mistral Small 3 Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1246#223 of 291, top 77%LMArena2026-10-08

Writing & Preference

Mistral Small 3 Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1234#228 of 297, top 77%LMArena2026-10-08
LMArena Creative Writing1195#236 of 295, top 80%LMArena2026-10-08
EQ-Bench Creative Writing707#110 of 115, top 96%EQ-Bench
LMArena Multi-Turn1217#233 of 295, top 79%LMArena2026-10-08

API pricing by provider

Mistral Small 3 API prices
RouteInput $/MOutput $/MCached input $/MChecked
openrouter$0.05$0.08—2026-10-10

Compare Mistral Small 3

Other Mistral AI models

Frequently asked questions

How good is Mistral Small 3?

Mistral Small 3 by Mistral AI ranks 278th of 354 ranked models on the Noometry Index as of October 2026, with a score of 31.2. Its strongest category is coding, where it ranks 207th. API pricing starts at $0.05 per million input tokens and $0.08 per million output tokens, with a 33K-token context window.

How much does Mistral Small 3 cost?

Mistral Small 3 costs $0.05 per million input tokens and $0.08 per million output tokens on openrouter.

What is Mistral Small 3's context window?

Mistral Small 3 accepts up to 33K tokens of input and can write up to 16K tokens in one response.

Is Mistral Small 3 open source?

Yes. Mistral Small 3's weights are downloadable from Hugging Face (mistralai/Mistral-Small-24B-Instruct-2501); check the license for commercial terms.

What are Mistral Small 3's strengths and weaknesses?

Relative to other ranked models, Mistral Small 3 places best in coding, long context, instruction following and lowest in math, writing & preference, knowledge.

What is Mistral Small 3 best at?

Its best category is coding, where it ranks 207th on Noometry.