Model comparison

C4ai Aya Expanse 32b vs Mercury

Mercury is the stronger model overall, scoring 37.6 to 35.9 on the Noometry Index.

Last verified . 8 shared benchmarks.

C4ai Aya Expanse 32b Cohere

35.9

Rank #221 Confirmed

Mercury Inception

37.6

Rank #199 Confirmed

Summary

  • They share 8 benchmarks with published results for both. C4ai Aya Expanse 32b scores higher in 1 category and Mercury in 5 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where C4ai Aya Expanse 32b leads 23.3 to 17.5.
  • C4ai Aya Expanse 32b has downloadable open weights; the other is API-only.

Side by side

C4ai Aya Expanse 32b and Mercury specifications
C4ai Aya Expanse 32bMercury
ProviderCohereInception
Noometry Index35.937.6
Released2024-10-24—
WeightsOpenProprietary
Context window128K—
Max output4K—
Input $ / M tokens——
Output $ / M tokens——
Results tracked189

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mercury leads

C4ai Aya Expanse 32b: 34.8 (#231), Mercury: 38.7 (#170)

Coding benchmarks
BenchmarkC4ai Aya Expanse 32bMercury
LMArena Coding11971322

Reasoning C4ai Aya Expanse 32b leads

C4ai Aya Expanse 32b: 23.3 (#180), Mercury: 17.5 (#293)

Reasoning benchmarks
BenchmarkC4ai Aya Expanse 32bMercury
LMArena Hard Prompts11931285
Kagi LLM Benchmark—21.6%

Math Not comparable

C4ai Aya Expanse 32b: 34.0 (#197), Mercury: —

Math benchmarks
BenchmarkC4ai Aya Expanse 32bMercury
LMArena Math1200—

Knowledge Not comparable

C4ai Aya Expanse 32b: 33.2 (#206), Mercury: —

Knowledge benchmarks
BenchmarkC4ai Aya Expanse 32bMercury
Vectara Hallucination Rate10.9%—
LMArena Expert1182—

Multilingual Mercury leads

C4ai Aya Expanse 32b: 38.4 (#230), Mercury: 41.6 (#206)

Multilingual benchmarks
BenchmarkC4ai Aya Expanse 32bMercury
LMArena Non-English12131260
LMArena Chinese1211—
LMArena French1248—
LMArena German1199—
LMArena Japanese1163—
LMArena Korean1158—
LMArena Russian1227—
LMArena Spanish1193—

Instruction Following Mercury leads

C4ai Aya Expanse 32b: 62.6 (#237), Mercury: 65.2 (#224)

Instruction Following benchmarks
BenchmarkC4ai Aya Expanse 32bMercury
LMArena Instruction Following11961239

Long Context Mercury leads

C4ai Aya Expanse 32b: 37.2 (#220), Mercury: 38.4 (#198)

Long Context benchmarks
BenchmarkC4ai Aya Expanse 32bMercury
LMArena Longer Query12281266

Writing & Preference Mercury leads

C4ai Aya Expanse 32b: 42.2 (#235), Mercury: 46.2 (#221)

Writing & Preference benchmarks
BenchmarkC4ai Aya Expanse 32bMercury
LMArena Text12241282
LMArena Creative Writing12001191
LMArena Multi-Turn11901282

Frequently asked questions

Is C4ai Aya Expanse 32b better than Mercury?

Mercury is the stronger model overall, scoring 37.6 to 35.9 on the Noometry Index.

Is C4ai Aya Expanse 32b or Mercury better for coding?

Mercury scores higher on coding benchmarks: 38.7 versus 34.8 in the Noometry coding category.

How many benchmarks do C4ai Aya Expanse 32b and Mercury share?

8 benchmarks have published results for both models. C4ai Aya Expanse 32b has 18 scored results on Noometry and Mercury has 9.

Related comparisons

Go deeper