Model comparison

C4ai Aya Expanse 32b vs Devstral Small 2505

C4ai Aya Expanse 32b is the stronger model overall, scoring 35.9 to 34.3 on the Noometry Index.

Last verified . 0 shared benchmarks.

C4ai Aya Expanse 32b Cohere

35.9

Rank #221 Confirmed

Devstral Small 2505 Mistral AI

34.3

Rank #233 Reported

Summary

  • The widest gap is in coding, where Devstral Small 2505 leads 38.9 to 34.8.

Side by side

C4ai Aya Expanse 32b and Devstral Small 2505 specifications
C4ai Aya Expanse 32bDevstral Small 2505
ProviderCohereMistral AI
Noometry Index35.934.3
Released2024-10-242025-05-07
WeightsOpenOpen
Context window128K128K
Max output4K128K
Input $ / M tokens—$0.10
Output $ / M tokens—$0.30
Results tracked184

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Devstral Small 2505 leads

C4ai Aya Expanse 32b: 34.8 (#231), Devstral Small 2505: 38.9 (#166)

Coding benchmarks
BenchmarkC4ai Aya Expanse 32bDevstral Small 2505
SWE-bench Verified (bash only)—56.4%
SciCode—28.8%
LMArena Coding1197—

Reasoning C4ai Aya Expanse 32b leads

C4ai Aya Expanse 32b: 23.3 (#180), Devstral Small 2505: 19.7 (#252)

Reasoning benchmarks
BenchmarkC4ai Aya Expanse 32bDevstral Small 2505
Kagi LLM Benchmark—37.7%
CritPt—0%
LMArena Hard Prompts1193—

Math Not comparable

C4ai Aya Expanse 32b: 34.0 (#197), Devstral Small 2505: —

Math benchmarks
BenchmarkC4ai Aya Expanse 32bDevstral Small 2505
LMArena Math1200—

Knowledge Not comparable

C4ai Aya Expanse 32b: 33.2 (#206), Devstral Small 2505: —

Knowledge benchmarks
BenchmarkC4ai Aya Expanse 32bDevstral Small 2505
Vectara Hallucination Rate10.9%—
LMArena Expert1182—

Multilingual Not comparable

C4ai Aya Expanse 32b: 38.4 (#230), Devstral Small 2505: —

Multilingual benchmarks
BenchmarkC4ai Aya Expanse 32bDevstral Small 2505
LMArena Non-English1213—
LMArena Chinese1211—
LMArena French1248—
LMArena German1199—
LMArena Japanese1163—
LMArena Korean1158—
LMArena Russian1227—
LMArena Spanish1193—

Instruction Following Not comparable

C4ai Aya Expanse 32b: 62.6 (#237), Devstral Small 2505: —

Instruction Following benchmarks
BenchmarkC4ai Aya Expanse 32bDevstral Small 2505
LMArena Instruction Following1196—

Long Context Not comparable

C4ai Aya Expanse 32b: 37.2 (#220), Devstral Small 2505: —

Long Context benchmarks
BenchmarkC4ai Aya Expanse 32bDevstral Small 2505
LMArena Longer Query1228—

Writing & Preference Not comparable

C4ai Aya Expanse 32b: 42.2 (#235), Devstral Small 2505: —

Writing & Preference benchmarks
BenchmarkC4ai Aya Expanse 32bDevstral Small 2505
LMArena Text1224—
LMArena Creative Writing1200—
LMArena Multi-Turn1190—

Frequently asked questions

Is C4ai Aya Expanse 32b better than Devstral Small 2505?

C4ai Aya Expanse 32b is the stronger model overall, scoring 35.9 to 34.3 on the Noometry Index.

Is C4ai Aya Expanse 32b or Devstral Small 2505 better for coding?

Devstral Small 2505 scores higher on coding benchmarks: 38.9 versus 34.8 in the Noometry coding category.

Which has the bigger context window?

Both accept 128K tokens.

How many benchmarks do C4ai Aya Expanse 32b and Devstral Small 2505 share?

0 benchmarks have published results for both models. C4ai Aya Expanse 32b has 18 scored results on Noometry and Devstral Small 2505 has 4.

Related comparisons

Go deeper