Model comparison

C4ai Aya Expanse 8b vs Mistral Small

C4ai Aya Expanse 8b is the stronger model overall, scoring 34.9 to 33.4 on the Noometry Index.

Last verified . 15 shared benchmarks.

C4ai Aya Expanse 8b Cohere

34.9

Rank #229 Confirmed

Mistral Small Mistral AI

33.4

Rank #243 Confirmed

Summary

  • They share 15 benchmarks with published results for both. C4ai Aya Expanse 8b scores higher in 3 categories and Mistral Small in 5 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in math, where C4ai Aya Expanse 8b leads 33.3 to 16.4.

Side by side

C4ai Aya Expanse 8b and Mistral Small specifications
C4ai Aya Expanse 8bMistral Small
ProviderCohereMistral AI
Noometry Index34.933.4
Released—2024-02-26
WeightsOpenOpen
Context window—262K
Max output—256K
Input $ / M tokens—$0.15
Output $ / M tokens—$0.60
Results tracked1539

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

C4ai Aya Expanse 8b: 33.8 (#252), Mistral Small: 34.0 (#247)

Coding benchmarks
BenchmarkC4ai Aya Expanse 8bMistral Small
LMArena Coding11601362
SciCode—26.5%
BigCodeBench Instruct—36.1%
LiveBench Coding—36.2%
BigCodeBench Complete—46.6%
ALE-Bench—497.62

Agentic & Tool Use Not comparable

C4ai Aya Expanse 8b: —, Mistral Small: 28.1 (#93)

Agentic & Tool Use benchmarks
BenchmarkC4ai Aya Expanse 8bMistral Small
Berkeley Function Calling Leaderboard—37.1%

Reasoning C4ai Aya Expanse 8b leads

C4ai Aya Expanse 8b: 22.4 (#197), Mistral Small: 19.8 (#250)

Reasoning benchmarks
BenchmarkC4ai Aya Expanse 8bMistral Small
LMArena Hard Prompts11561335
Kagi LLM Benchmark—37.8%
CritPt—0%
LiveBench Reasoning—44.8%
DTBench—70.9%
LiveBench Data Analysis—53.7%
LMCA—20.6%
LiveBench—44%

Math C4ai Aya Expanse 8b leads

C4ai Aya Expanse 8b: 33.3 (#203), Mistral Small: 16.4 (#293)

Math benchmarks
BenchmarkC4ai Aya Expanse 8bMistral Small
LMArena Math11681341
OTIS Mock AIME 2024-2025—5.8%
LiveBench Math—39.9%
MATH Level 5—46.8%

Knowledge C4ai Aya Expanse 8b leads

C4ai Aya Expanse 8b: 33.6 (#201), Mistral Small: 31.0 (#222)

Knowledge benchmarks
BenchmarkC4ai Aya Expanse 8bMistral Small
Vectara Hallucination Rate9.5%5.1%
LMArena Expert11531291
GPQA Diamond—47.5%
MMLU—68.7%

Multimodal Not comparable

C4ai Aya Expanse 8b: —, Mistral Small: 33.5 (#96)

Multimodal benchmarks
BenchmarkC4ai Aya Expanse 8bMistral Small
LMArena Vision—1142

Multilingual Mistral Small leads

C4ai Aya Expanse 8b: 36.1 (#242), Mistral Small: 45.5 (#169)

Multilingual benchmarks
BenchmarkC4ai Aya Expanse 8bMistral Small
LMArena Non-English11801315
LMArena Chinese11811340
LMArena German11881340
LMArena Japanese11201275
LMArena Russian11971324
LMArena French—1337
LMArena Korean—1259
LMArena Spanish—1346

Instruction Following Mistral Small leads

C4ai Aya Expanse 8b: 60.1 (#253), Mistral Small: 66.4 (#209)

Instruction Following benchmarks
BenchmarkC4ai Aya Expanse 8bMistral Small
LMArena Instruction Following11551310
LiveBench Instruction Following—63.7%

Long Context Mistral Small leads

C4ai Aya Expanse 8b: 36.1 (#236), Mistral Small: 40.4 (#156)

Long Context benchmarks
BenchmarkC4ai Aya Expanse 8bMistral Small
LMArena Longer Query11911327

Writing & Preference Mistral Small leads

C4ai Aya Expanse 8b: 39.0 (#250), Mistral Small: 52.5 (#171)

Writing & Preference benchmarks
BenchmarkC4ai Aya Expanse 8bMistral Small
LMArena Text11851338
LMArena Creative Writing11681305
LMArena Multi-Turn11601344
LiveBench Language—30.5%

Frequently asked questions

Is C4ai Aya Expanse 8b better than Mistral Small?

C4ai Aya Expanse 8b is the stronger model overall, scoring 34.9 to 33.4 on the Noometry Index.

Is C4ai Aya Expanse 8b or Mistral Small better for coding?

They score almost the same on coding (33.8 vs 34.0); test both on your own repository before choosing.

How many benchmarks do C4ai Aya Expanse 8b and Mistral Small share?

15 benchmarks have published results for both models. C4ai Aya Expanse 8b has 15 scored results on Noometry and Mistral Small has 39.

Related comparisons

Go deeper