Model comparison

C4ai Aya Expanse 32b vs Mistral Small 3

C4ai Aya Expanse 32b is the stronger model overall, scoring 35.9 to 31.2 on the Noometry Index.

Last verified . 16 shared benchmarks.

C4ai Aya Expanse 32b Cohere

35.9

Rank #221 Confirmed

Mistral Small 3 Mistral AI

31.2

Rank #278 Confirmed

Summary

  • They share 16 benchmarks with published results for both. C4ai Aya Expanse 32b scores higher in 5 categories and Mistral Small 3 in 3 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in math, where C4ai Aya Expanse 32b leads 34.0 to 16.3.
  • C4ai Aya Expanse 32b accepts more context: 128K tokens versus 33K.

Side by side

C4ai Aya Expanse 32b and Mistral Small 3 specifications
C4ai Aya Expanse 32bMistral Small 3
ProviderCohereMistral AI
Noometry Index35.931.2
Released2024-10-242025-01-30
WeightsOpenOpen
Context window128K33K
Max output4K16K
Input $ / M tokens—$0.05
Output $ / M tokens—$0.08
Results tracked1824

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Small 3 leads

C4ai Aya Expanse 32b: 34.8 (#231), Mistral Small 3: 36.5 (#207)

Coding benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Small 3
LMArena Coding11971246
BigCodeBench Instruct—45.3%
BigCodeBench Complete—50.4%

Reasoning C4ai Aya Expanse 32b leads

C4ai Aya Expanse 32b: 23.3 (#180), Mistral Small 3: 18.9 (#273)

Reasoning benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Small 3
LMArena Hard Prompts11931233
Chess Puzzles—0%
Epoch Capabilities Index—127.07

Math C4ai Aya Expanse 32b leads

C4ai Aya Expanse 32b: 34.0 (#197), Mistral Small 3: 16.3 (#295)

Math benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Small 3
LMArena Math12001240
OTIS Mock AIME 2024-2025—6.7%

Knowledge C4ai Aya Expanse 32b leads

C4ai Aya Expanse 32b: 33.2 (#206), Mistral Small 3: 25.1 (#263)

Knowledge benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Small 3
LMArena Expert11821202
GPQA Diamond—47.3%
Confabulations—25.2%
Vectara Hallucination Rate10.9%—

Multilingual C4ai Aya Expanse 32b leads

C4ai Aya Expanse 32b: 38.4 (#230), Mistral Small 3: 37.3 (#236)

Multilingual benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Small 3
LMArena Non-English12131198
LMArena Chinese12111204
LMArena French12481203
LMArena German11991211
LMArena Japanese11631111
LMArena Korean11581188
LMArena Russian12271216
LMArena Spanish1193—

Instruction Following Mistral Small 3 leads

C4ai Aya Expanse 32b: 62.6 (#237), Mistral Small 3: 63.7 (#229)

Instruction Following benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Small 3
LMArena Instruction Following11961214

Long Context Too close to call

C4ai Aya Expanse 32b: 37.2 (#220), Mistral Small 3: 37.8 (#211)

Long Context benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Small 3
LMArena Longer Query12281246

Writing & Preference C4ai Aya Expanse 32b leads

C4ai Aya Expanse 32b: 42.2 (#235), Mistral Small 3: 32.2 (#280)

Writing & Preference benchmarks
BenchmarkC4ai Aya Expanse 32bMistral Small 3
LMArena Text12241234
LMArena Creative Writing12001195
LMArena Multi-Turn11901217
EQ-Bench Creative Writing—707

Frequently asked questions

Is C4ai Aya Expanse 32b better than Mistral Small 3?

C4ai Aya Expanse 32b is the stronger model overall, scoring 35.9 to 31.2 on the Noometry Index.

Is C4ai Aya Expanse 32b or Mistral Small 3 better for coding?

Mistral Small 3 scores higher on coding benchmarks: 36.5 versus 34.8 in the Noometry coding category.

Which has the bigger context window?

C4ai Aya Expanse 32b does, with 128K tokens against 33K.

How many benchmarks do C4ai Aya Expanse 32b and Mistral Small 3 share?

16 benchmarks have published results for both models. C4ai Aya Expanse 32b has 18 scored results on Noometry and Mistral Small 3 has 24.

Related comparisons

Go deeper