Model comparison

Mistral Medium vs Nova 2 Lite

Nova 2 Lite is the stronger model overall, scoring 39.7 to 36.3 on the Noometry Index.

Last verified . 19 shared benchmarks.

Mistral Medium Mistral AI

36.3

Rank #218 Confirmed

Nova 2 Lite Amazon

39.7

Rank #161 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Mistral Medium scores higher in 5 categories and Nova 2 Lite in 4 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Nova 2 Lite leads 43.0 to 25.0.
  • The biggest single-benchmark swing is Vectara Hallucination Rate: 22.7% for Mistral Medium and 5.1% for Nova 2 Lite.
  • Nova 2 Lite is cheaper at $0.30 / $2.50 per million input/output tokens, against $1.50 / $7.50 for Mistral Medium.
  • Nova 2 Lite accepts more context: 1M tokens versus 262K.
  • Mistral Medium has downloadable open weights; the other is API-only.

Side by side

Mistral Medium and Nova 2 Lite specifications
Mistral MediumNova 2 Lite
ProviderMistral AIAmazon
Noometry Index36.339.7
Released2023-12-112025-12-01
WeightsOpenProprietary
Context window262K1M
Max output262K64K
Input $ / M tokens$1.50$0.30
Output $ / M tokens$7.50$2.50
Results tracked3619

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Nova 2 Lite leads

Mistral Medium: 34.2 (#243), Nova 2 Lite: 40.7 (#134)

Coding benchmarks
BenchmarkMistral MediumNova 2 Lite
LMArena Coding14341385
FrontierCode8%—
SciCode40.2%—
WeirdML43.7%—
ALE-Bench763.98—

Agentic & Tool Use Mistral Medium leads

Mistral Medium: 28.3 (#90), Nova 2 Lite: 24.1 (#121)

Agentic & Tool Use benchmarks
BenchmarkMistral MediumNova 2 Lite
Berkeley Function Calling Leaderboard37.7%27.1%

Reasoning Nova 2 Lite leads

Mistral Medium: 24.0 (#167), Nova 2 Lite: 27.5 (#118)

Reasoning benchmarks
BenchmarkMistral MediumNova 2 Lite
LMArena Hard Prompts14261364
Kagi LLM Benchmark50%—
CritPt0%—
DTBench75.5%—
LMCA26.1%—
Surface Evolver Bench26.9%—

Math Nova 2 Lite leads

Mistral Medium: 28.1 (#245), Nova 2 Lite: 37.5 (#156)

Math benchmarks
BenchmarkMistral MediumNova 2 Lite
LMArena Math14081359
OTIS Mock AIME 2024-202532.2%—
ProofBench9%—
MATH Level 581.6%—
FrontierMath (Feb 2025 set)0.3%—

Knowledge Nova 2 Lite leads

Mistral Medium: 25.0 (#265), Nova 2 Lite: 43.0 (#94)

Knowledge benchmarks
BenchmarkMistral MediumNova 2 Lite
Vectara Hallucination Rate22.7%5.1%
LMArena Expert14081358
GPQA Diamond59.5%—
Humanity's Last Exam4.5%—

Multimodal Not comparable

Mistral Medium: 35.3 (#88), Nova 2 Lite: —

Multimodal benchmarks
BenchmarkMistral MediumNova 2 Lite
LMArena Vision1172—

Multilingual Mistral Medium leads

Mistral Medium: 52.1 (#91), Nova 2 Lite: 47.1 (#153)

Multilingual benchmarks
BenchmarkMistral MediumNova 2 Lite
LMArena Non-English14081337
LMArena Chinese14471364
LMArena French14591381
LMArena German14321343
LMArena Japanese13781271
LMArena Korean13801284
LMArena Russian14111343
LMArena Spanish14331373

Instruction Following Mistral Medium leads

Mistral Medium: 73.7 (#116), Nova 2 Lite: 70.5 (#161)

Instruction Following benchmarks
BenchmarkMistral MediumNova 2 Lite
LMArena Instruction Following13981335

Long Context Mistral Medium leads

Mistral Medium: 42.9 (#114), Nova 2 Lite: 40.6 (#150)

Long Context benchmarks
BenchmarkMistral MediumNova 2 Lite
LMArena Longer Query14061335

Writing & Preference Mistral Medium leads

Mistral Medium: 60.0 (#103), Nova 2 Lite: 53.9 (#154)

Writing & Preference benchmarks
BenchmarkMistral MediumNova 2 Lite
LMArena Text14241362
LMArena Creative Writing13911291
LMArena Multi-Turn14181338
Short-Story Creative Writing77.3%—

Frequently asked questions

Is Mistral Medium better than Nova 2 Lite?

Nova 2 Lite is the stronger model overall, scoring 39.7 to 36.3 on the Noometry Index.

Which is cheaper, Mistral Medium or Nova 2 Lite?

Nova 2 Lite is cheaper. It lists at $0.30 per million input tokens and $2.50 per million output tokens; Mistral Medium lists at $1.50 and $7.50.

Is Mistral Medium or Nova 2 Lite better for coding?

Nova 2 Lite scores higher on coding benchmarks: 40.7 versus 34.2 in the Noometry coding category.

Which has the bigger context window?

Nova 2 Lite does, with 1M tokens against 262K.

How many benchmarks do Mistral Medium and Nova 2 Lite share?

19 benchmarks have published results for both models. Mistral Medium has 36 scored results on Noometry and Nova 2 Lite has 19.

Related comparisons

Go deeper