Model comparison

Amazon Nova Experimental Chat 26 01 10 vs Mistral Small 3.2

Amazon Nova Experimental Chat 26 01 10 is the stronger model overall, scoring 42.8 to 31.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Mistral Small 3.2 Mistral AI

31.2

Rank #280 Confirmed

Summary

  • The widest gap is in knowledge, where Amazon Nova Experimental Chat 26 01 10 leads 40.3 to 26.7.
  • Mistral Small 3.2 has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Experimental Chat 26 01 10 and Mistral Small 3.2 specifications
Amazon Nova Experimental Chat 26 01 10Mistral Small 3.2
ProviderAmazonMistral AI
Noometry Index42.831.2
Released—2025-06-20
WeightsProprietaryOpen
Context window—256K
Max output—16K
Input $ / M tokens—$0.0938
Output $ / M tokens—$0.25
Results tracked136

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Amazon Nova Experimental Chat 26 01 10: 42.7 (#96), Mistral Small 3.2: —

Coding benchmarks
BenchmarkAmazon Nova Experimental Chat 26 01 10Mistral Small 3.2
LMArena Coding1447—

Reasoning Amazon Nova Experimental Chat 26 01 10 leads

Amazon Nova Experimental Chat 26 01 10: 28.9 (#98), Mistral Small 3.2: 18.1 (#287)

Reasoning benchmarks
BenchmarkAmazon Nova Experimental Chat 26 01 10Mistral Small 3.2
Kagi LLM Benchmark—40.4%
Chess Puzzles—1%
LMArena Hard Prompts1415—
Epoch Capabilities Index—131.74

Math Amazon Nova Experimental Chat 26 01 10 leads

Amazon Nova Experimental Chat 26 01 10: 38.5 (#136), Mistral Small 3.2: 26.3 (#260)

Math benchmarks
BenchmarkAmazon Nova Experimental Chat 26 01 10Mistral Small 3.2
OTIS Mock AIME 2024-2025—30.3%
LMArena Math1404—

Knowledge Amazon Nova Experimental Chat 26 01 10 leads

Amazon Nova Experimental Chat 26 01 10: 40.3 (#120), Mistral Small 3.2: 26.7 (#256)

Knowledge benchmarks
BenchmarkAmazon Nova Experimental Chat 26 01 10Mistral Small 3.2
GPQA Diamond—49.1%
LMArena Expert1444—

Multilingual Not comparable

Amazon Nova Experimental Chat 26 01 10: 50.2 (#126), Mistral Small 3.2: —

Multilingual benchmarks
BenchmarkAmazon Nova Experimental Chat 26 01 10Mistral Small 3.2
LMArena Non-English1381—
LMArena Chinese1372—
LMArena Russian1391—
LMArena Spanish1390—

Instruction Following Not comparable

Amazon Nova Experimental Chat 26 01 10: 72.9 (#126), Mistral Small 3.2: —

Instruction Following benchmarks
BenchmarkAmazon Nova Experimental Chat 26 01 10Mistral Small 3.2
LMArena Instruction Following1381—

Long Context Not comparable

Amazon Nova Experimental Chat 26 01 10: 42.7 (#119), Mistral Small 3.2: —

Long Context benchmarks
BenchmarkAmazon Nova Experimental Chat 26 01 10Mistral Small 3.2
LMArena Longer Query1400—

Writing & Preference Amazon Nova Experimental Chat 26 01 10 leads

Amazon Nova Experimental Chat 26 01 10: 57.4 (#129), Mistral Small 3.2: 45.0 (#224)

Writing & Preference benchmarks
BenchmarkAmazon Nova Experimental Chat 26 01 10Mistral Small 3.2
LMArena Text1396—
LMArena Creative Writing1331—
EQ-Bench Creative Writing—1255
LMArena Multi-Turn1380—

Frequently asked questions

Is Amazon Nova Experimental Chat 26 01 10 better than Mistral Small 3.2?

Amazon Nova Experimental Chat 26 01 10 is the stronger model overall, scoring 42.8 to 31.2 on the Noometry Index.

How many benchmarks do Amazon Nova Experimental Chat 26 01 10 and Mistral Small 3.2 share?

0 benchmarks have published results for both models. Amazon Nova Experimental Chat 26 01 10 has 13 scored results on Noometry and Mistral Small 3.2 has 6.

Related comparisons

Go deeper