Model comparison

Mistral Small 3 vs Palm 2

Mistral Small 3 is the stronger model overall, scoring 31.2 to 30.0 on the Noometry Index.

Last verified . 10 shared benchmarks.

Mistral Small 3 Mistral AI

31.2

Rank #278 Confirmed

Palm 2 Google

30.0

Rank #298 Confirmed

Summary

  • They share 10 benchmarks with published results for both. Mistral Small 3 scores higher in 5 categories and Palm 2 in 2 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in multilingual, where Mistral Small 3 leads 37.3 to 20.9.
  • Mistral Small 3 has downloadable open weights; the other is API-only.

Side by side

Mistral Small 3 and Palm 2 specifications
Mistral Small 3Palm 2
ProviderMistral AIGoogle
Noometry Index31.230.0
Released2025-01-30—
WeightsOpenProprietary
Context window33K—
Max output16K—
Input $ / M tokens$0.05—
Output $ / M tokens$0.08—
Results tracked2410

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Small 3 leads

Mistral Small 3: 36.5 (#207), Palm 2: 29.0 (#311)

Coding benchmarks
BenchmarkMistral Small 3Palm 2
LMArena Coding1246994
BigCodeBench Instruct45.3%—
BigCodeBench Complete50.4%—

Reasoning Too close to call

Mistral Small 3: 18.9 (#273), Palm 2: 19.1 (#271)

Reasoning benchmarks
BenchmarkMistral Small 3Palm 2
LMArena Hard Prompts12331005
Chess Puzzles0%—
Epoch Capabilities Index127.07—

Math Palm 2 leads

Mistral Small 3: 16.3 (#295), Palm 2: 30.8 (#231)

Math benchmarks
BenchmarkMistral Small 3Palm 2
LMArena Math12401049
OTIS Mock AIME 2024-20256.7%—

Knowledge Not comparable

Mistral Small 3: 25.1 (#263), Palm 2: —

Knowledge benchmarks
BenchmarkMistral Small 3Palm 2
GPQA Diamond47.3%—
Confabulations25.2%—
LMArena Expert1202—

Multilingual Mistral Small 3 leads

Mistral Small 3: 37.3 (#236), Palm 2: 20.9 (#295)

Multilingual benchmarks
BenchmarkMistral Small 3Palm 2
LMArena Non-English1198916
LMArena Chinese1204887
LMArena French1203—
LMArena German1211—
LMArena Japanese1111—
LMArena Korean1188—
LMArena Russian1216—

Instruction Following Mistral Small 3 leads

Mistral Small 3: 63.7 (#229), Palm 2: 51.0 (#296)

Instruction Following benchmarks
BenchmarkMistral Small 3Palm 2
LMArena Instruction Following12141010

Long Context Mistral Small 3 leads

Mistral Small 3: 37.8 (#211), Palm 2: 30.8 (#285)

Long Context benchmarks
BenchmarkMistral Small 3Palm 2
LMArena Longer Query12461010

Writing & Preference Mistral Small 3 leads

Mistral Small 3: 32.2 (#280), Palm 2: 25.6 (#304)

Writing & Preference benchmarks
BenchmarkMistral Small 3Palm 2
LMArena Text12341027
LMArena Creative Writing1195991
LMArena Multi-Turn1217994
EQ-Bench Creative Writing707—

Frequently asked questions

Is Mistral Small 3 better than Palm 2?

Mistral Small 3 is the stronger model overall, scoring 31.2 to 30.0 on the Noometry Index.

Is Mistral Small 3 or Palm 2 better for coding?

Mistral Small 3 scores higher on coding benchmarks: 36.5 versus 29.0 in the Noometry coding category.

How many benchmarks do Mistral Small 3 and Palm 2 share?

10 benchmarks have published results for both models. Mistral Small 3 has 24 scored results on Noometry and Palm 2 has 10.

Related comparisons

Go deeper