Model comparison

GPT-5 Nano vs Mistral

GPT-5 Nano is the stronger model overall, scoring 33.5 to 29.9 on the Noometry Index.

Last verified . 21 shared benchmarks.

GPT-5 Nano OpenAI

33.5

Rank #241 Confirmed

Mistral Mistral AI

29.9

Rank #303 Confirmed

Summary

  • They share 21 benchmarks with published results for both. GPT-5 Nano scores higher in 5 categories and Mistral in 3 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in instruction following, where GPT-5 Nano leads 75.0 to 52.6.
  • The biggest single-benchmark swing is MMLU-Pro: 77.8% for GPT-5 Nano and 27.7% for Mistral.

Side by side

GPT-5 Nano and Mistral specifications
GPT-5 NanoMistral
ProviderOpenAIMistral AI
Noometry Index33.529.9
Released2025-08-07—
WeightsProprietaryProprietary
Context window400K—
Max output128K—
Input $ / M tokens$0.05—
Output $ / M tokens$0.40—
Results tracked4922

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

GPT-5 Nano: 33.6 (#254), Mistral: 33.8 (#250)

Coding benchmarks
BenchmarkGPT-5 NanoMistral
LMArena Coding13511162
SWE-bench Verified (bash only)34.8%—
WeirdML38.1%—
ALE-Bench718.67—

Agentic & Tool Use Not comparable

GPT-5 Nano: 25.8 (#106), Mistral: —

Agentic & Tool Use benchmarks
BenchmarkGPT-5 NanoMistral
Terminal-Bench21.8%—
Berkeley Function Calling Leaderboard51.5%—

Reasoning Mistral leads

GPT-5 Nano: 16.3 (#306), Mistral: 22.2 (#200)

Reasoning benchmarks
BenchmarkGPT-5 NanoMistral
LMArena Hard Prompts13281149
ARC-AGI-22.6%—
Kagi LLM Benchmark62.2%—
ARC-AGI-120.7%—
Chess Puzzles27%—
Mystery Game Puzzles9%—
DTBench62.7%—
LMCA7.9%—
Epoch Capabilities Index139.38—
ForecastBench59.1—

Math GPT-5 Nano leads

GPT-5 Nano: 29.4 (#241), Mistral: 22.3 (#278)

Knowledge GPT-5 Nano leads

GPT-5 Nano: 35.9 (#178), Mistral: 16.6 (#288)

Knowledge benchmarks
BenchmarkGPT-5 NanoMistral
MMLU-Pro77.8%27.7%
GPQA (HELM)67.9%30.3%
LMArena Expert13211125
GPQA Diamond69.4%—
SimpleQA Verified11.7%—
Vectara Hallucination Rate10.5%—

Multimodal Not comparable

GPT-5 Nano: 31.3 (#108), Mistral: —

Multimodal benchmarks
BenchmarkGPT-5 NanoMistral
LMArena Vision1159—
VPCT37.2%—

Multilingual GPT-5 Nano leads

GPT-5 Nano: 45.3 (#172), Mistral: 32.8 (#254)

Multilingual benchmarks
BenchmarkGPT-5 NanoMistral
LMArena Non-English13131129
LMArena Chinese13561109
LMArena German13271155
LMArena Japanese12261013
LMArena Korean12691032
LMArena Russian12961168
LMArena Spanish13601143
LMArena French—1180

Instruction Following GPT-5 Nano leads

GPT-5 Nano: 75.0 (#79), Mistral: 52.6 (#288)

Instruction Following benchmarks
BenchmarkGPT-5 NanoMistral
IFEval93.2%56.8%
LMArena Instruction Following13061152

Long Context Mistral leads

GPT-5 Nano: 31.3 (#281), Mistral: 35.0 (#245)

Long Context benchmarks
BenchmarkGPT-5 NanoMistral
LMArena Longer Query13121153
Fiction.LiveBench44.4%—

Writing & Preference GPT-5 Nano leads

GPT-5 Nano: 39.1 (#249), Mistral: 37.0 (#260)

Writing & Preference benchmarks
BenchmarkGPT-5 NanoMistral
LMArena Text13201165
LMArena Creative Writing12491158
WildBench80.6%66%
LMArena Multi-Turn13111147
EQ-Bench Creative Writing705—

Frequently asked questions

Is GPT-5 Nano better than Mistral?

GPT-5 Nano is the stronger model overall, scoring 33.5 to 29.9 on the Noometry Index.

Is GPT-5 Nano or Mistral better for coding?

They score almost the same on coding (33.6 vs 33.8); test both on your own repository before choosing.

How many benchmarks do GPT-5 Nano and Mistral share?

21 benchmarks have published results for both models. GPT-5 Nano has 49 scored results on Noometry and Mistral has 22.

Related comparisons

Go deeper