Model comparison

Mistral Medium vs Nemotron 3.5 Lightning

Nemotron 3.5 Lightning is the stronger model overall, scoring 40.0 to 36.3 on the Noometry Index.

Last verified . 17 shared benchmarks.

Mistral Medium Mistral AI

36.3

Rank #218 Confirmed

Nemotron 3.5 Lightning NVIDIA

40.0

Rank #155 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Mistral Medium scores higher in 4 categories and Nemotron 3.5 Lightning in 4 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Nemotron 3.5 Lightning leads 37.5 to 25.0.
  • Nemotron 3.5 Lightning is cheaper at $0.05 / $0.20 per million input/output tokens, against $1.50 / $7.50 for Mistral Medium.

Side by side

Mistral Medium and Nemotron 3.5 Lightning specifications
Mistral MediumNemotron 3.5 Lightning
ProviderMistral AINVIDIA
Noometry Index36.340.0
Released2023-12-112026-08-11
WeightsOpenOpen
Context window262K262K
Max output262K262K
Input $ / M tokens$1.50$0.05
Output $ / M tokens$7.50$0.20
Results tracked3618

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Nemotron 3.5 Lightning leads

Mistral Medium: 34.2 (#243), Nemotron 3.5 Lightning: 40.4 (#141)

Coding benchmarks
BenchmarkMistral MediumNemotron 3.5 Lightning
LMArena Coding14341375
FrontierCode8%—
SciCode40.2%—
WeirdML43.7%—
ALE-Bench763.98—

Agentic & Tool Use Not comparable

Mistral Medium: 28.3 (#90), Nemotron 3.5 Lightning: —

Agentic & Tool Use benchmarks
BenchmarkMistral MediumNemotron 3.5 Lightning
Berkeley Function Calling Leaderboard37.7%—

Reasoning Nemotron 3.5 Lightning leads

Mistral Medium: 24.0 (#167), Nemotron 3.5 Lightning: 26.8 (#127)

Reasoning benchmarks
BenchmarkMistral MediumNemotron 3.5 Lightning
LMArena Hard Prompts14261337
Kagi LLM Benchmark50%—
CritPt0%—
DTBench75.5%—
LMCA26.1%—
Surface Evolver Bench26.9%—

Math Nemotron 3.5 Lightning leads

Mistral Medium: 28.1 (#245), Nemotron 3.5 Lightning: 37.5 (#155)

Math benchmarks
BenchmarkMistral MediumNemotron 3.5 Lightning
LMArena Math14081359
OTIS Mock AIME 2024-202532.2%—
ProofBench9%—
MATH Level 581.6%—
FrontierMath (Feb 2025 set)0.3%—

Knowledge Nemotron 3.5 Lightning leads

Mistral Medium: 25.0 (#265), Nemotron 3.5 Lightning: 37.5 (#154)

Knowledge benchmarks
BenchmarkMistral MediumNemotron 3.5 Lightning
LMArena Expert14081356
GPQA Diamond59.5%—
Humanity's Last Exam4.5%—
Vectara Hallucination Rate22.7%—

Multimodal Not comparable

Mistral Medium: 35.3 (#88), Nemotron 3.5 Lightning: —

Multimodal benchmarks
BenchmarkMistral MediumNemotron 3.5 Lightning
LMArena Vision1172—

Multilingual Mistral Medium leads

Mistral Medium: 52.1 (#91), Nemotron 3.5 Lightning: 44.0 (#180)

Multilingual benchmarks
BenchmarkMistral MediumNemotron 3.5 Lightning
LMArena Non-English14081295
LMArena Chinese14471359
LMArena French14591366
LMArena German14321282
LMArena Japanese13781206
LMArena Korean13801238
LMArena Russian14111253
LMArena Spanish14331345

Instruction Following Mistral Medium leads

Mistral Medium: 73.7 (#116), Nemotron 3.5 Lightning: 69.6 (#170)

Instruction Following benchmarks
BenchmarkMistral MediumNemotron 3.5 Lightning
LMArena Instruction Following13981318

Long Context Mistral Medium leads

Mistral Medium: 42.9 (#114), Nemotron 3.5 Lightning: 39.9 (#165)

Long Context benchmarks
BenchmarkMistral MediumNemotron 3.5 Lightning
LMArena Longer Query14061314

Writing & Preference Mistral Medium leads

Mistral Medium: 60.0 (#103), Nemotron 3.5 Lightning: 48.5 (#201)

Writing & Preference benchmarks
BenchmarkMistral MediumNemotron 3.5 Lightning
LMArena Text14241327
LMArena Creative Writing13911254
LMArena Multi-Turn14181328
Short-Story Creative Writing77.3%—
EQ-Bench Creative Writing—1280

Frequently asked questions

Is Mistral Medium better than Nemotron 3.5 Lightning?

Nemotron 3.5 Lightning is the stronger model overall, scoring 40.0 to 36.3 on the Noometry Index.

Which is cheaper, Mistral Medium or Nemotron 3.5 Lightning?

Nemotron 3.5 Lightning is cheaper. It lists at $0.05 per million input tokens and $0.20 per million output tokens; Mistral Medium lists at $1.50 and $7.50.

Is Mistral Medium or Nemotron 3.5 Lightning better for coding?

Nemotron 3.5 Lightning scores higher on coding benchmarks: 40.4 versus 34.2 in the Noometry coding category.

Which has the bigger context window?

Both accept 262K tokens.

How many benchmarks do Mistral Medium and Nemotron 3.5 Lightning share?

17 benchmarks have published results for both models. Mistral Medium has 36 scored results on Noometry and Nemotron 3.5 Lightning has 18.

Related comparisons

Go deeper