Model comparison

Mistral Small vs Nemotron 3.5 Lightning

Nemotron 3.5 Lightning is the stronger model overall, scoring 40.0 to 33.4 on the Noometry Index.

Last verified . 17 shared benchmarks.

Mistral Small Mistral AI

33.4

Rank #243 Confirmed

Nemotron 3.5 Lightning NVIDIA

40.0

Rank #155 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Mistral Small scores higher in 3 categories and Nemotron 3.5 Lightning in 5 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in math, where Nemotron 3.5 Lightning leads 37.5 to 16.4.
  • Nemotron 3.5 Lightning is cheaper at $0.05 / $0.20 per million input/output tokens, against $0.15 / $0.60 for Mistral Small.

Side by side

Mistral Small and Nemotron 3.5 Lightning specifications
Mistral SmallNemotron 3.5 Lightning
ProviderMistral AINVIDIA
Noometry Index33.440.0
Released2024-02-262026-08-11
WeightsOpenOpen
Context window262K262K
Max output256K262K
Input $ / M tokens$0.15$0.05
Output $ / M tokens$0.60$0.20
Results tracked3918

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Nemotron 3.5 Lightning leads

Mistral Small: 34.0 (#247), Nemotron 3.5 Lightning: 40.4 (#141)

Coding benchmarks
BenchmarkMistral SmallNemotron 3.5 Lightning
LMArena Coding13621375
SciCode26.5%—
BigCodeBench Instruct36.1%—
LiveBench Coding36.2%—
BigCodeBench Complete46.6%—
ALE-Bench497.62—

Agentic & Tool Use Not comparable

Mistral Small: 28.1 (#93), Nemotron 3.5 Lightning: —

Agentic & Tool Use benchmarks
BenchmarkMistral SmallNemotron 3.5 Lightning
Berkeley Function Calling Leaderboard37.1%—

Reasoning Nemotron 3.5 Lightning leads

Mistral Small: 19.8 (#250), Nemotron 3.5 Lightning: 26.8 (#127)

Reasoning benchmarks
BenchmarkMistral SmallNemotron 3.5 Lightning
LMArena Hard Prompts13351337
Kagi LLM Benchmark37.8%—
CritPt0%—
LiveBench Reasoning44.8%—
DTBench70.9%—
LiveBench Data Analysis53.7%—
LMCA20.6%—
LiveBench44%—

Math Nemotron 3.5 Lightning leads

Mistral Small: 16.4 (#293), Nemotron 3.5 Lightning: 37.5 (#155)

Math benchmarks
BenchmarkMistral SmallNemotron 3.5 Lightning
LMArena Math13411359
OTIS Mock AIME 2024-20255.8%—
LiveBench Math39.9%—
MATH Level 546.8%—

Knowledge Nemotron 3.5 Lightning leads

Mistral Small: 31.0 (#222), Nemotron 3.5 Lightning: 37.5 (#154)

Knowledge benchmarks
BenchmarkMistral SmallNemotron 3.5 Lightning
LMArena Expert12911356
GPQA Diamond47.5%—
Vectara Hallucination Rate5.1%—
MMLU68.7%—

Multimodal Not comparable

Mistral Small: 33.5 (#96), Nemotron 3.5 Lightning: —

Multimodal benchmarks
BenchmarkMistral SmallNemotron 3.5 Lightning
LMArena Vision1142—

Multilingual Mistral Small leads

Mistral Small: 45.5 (#169), Nemotron 3.5 Lightning: 44.0 (#180)

Multilingual benchmarks
BenchmarkMistral SmallNemotron 3.5 Lightning
LMArena Non-English13151295
LMArena Chinese13401359
LMArena French13371366
LMArena German13401282
LMArena Japanese12751206
LMArena Korean12591238
LMArena Russian13241253
LMArena Spanish13461345

Instruction Following Nemotron 3.5 Lightning leads

Mistral Small: 66.4 (#209), Nemotron 3.5 Lightning: 69.6 (#170)

Instruction Following benchmarks
BenchmarkMistral SmallNemotron 3.5 Lightning
LMArena Instruction Following13101318
LiveBench Instruction Following63.7%—

Long Context Too close to call

Mistral Small: 40.4 (#156), Nemotron 3.5 Lightning: 39.9 (#165)

Long Context benchmarks
BenchmarkMistral SmallNemotron 3.5 Lightning
LMArena Longer Query13271314

Writing & Preference Mistral Small leads

Mistral Small: 52.5 (#171), Nemotron 3.5 Lightning: 48.5 (#201)

Writing & Preference benchmarks
BenchmarkMistral SmallNemotron 3.5 Lightning
LMArena Text13381327
LMArena Creative Writing13051254
LMArena Multi-Turn13441328
EQ-Bench Creative Writing—1280
LiveBench Language30.5%—

Frequently asked questions

Is Mistral Small better than Nemotron 3.5 Lightning?

Nemotron 3.5 Lightning is the stronger model overall, scoring 40.0 to 33.4 on the Noometry Index.

Which is cheaper, Mistral Small or Nemotron 3.5 Lightning?

Nemotron 3.5 Lightning is cheaper. It lists at $0.05 per million input tokens and $0.20 per million output tokens; Mistral Small lists at $0.15 and $0.60.

Is Mistral Small or Nemotron 3.5 Lightning better for coding?

Nemotron 3.5 Lightning scores higher on coding benchmarks: 40.4 versus 34.0 in the Noometry coding category.

Which has the bigger context window?

Both accept 262K tokens.

How many benchmarks do Mistral Small and Nemotron 3.5 Lightning share?

17 benchmarks have published results for both models. Mistral Small has 39 scored results on Noometry and Nemotron 3.5 Lightning has 18.

Related comparisons

Go deeper