Model comparison

C4ai Aya Expanse 32b vs Gemini 3.5 Flash

Gemini 3.5 Flash is the stronger model overall, scoring 54.2 to 35.9 on the Noometry Index.

Last verified . 17 shared benchmarks.

C4ai Aya Expanse 32b Cohere

35.9

Rank #221 Confirmed

Gemini 3.5 Flash Google

54.2

Rank #32 Confirmed

Summary

  • They share 17 benchmarks with published results for both. C4ai Aya Expanse 32b scores higher in 0 categories and Gemini 3.5 Flash in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Gemini 3.5 Flash leads 62.8 to 23.3.
  • Gemini 3.5 Flash accepts more context: 1.05M tokens versus 128K.
  • C4ai Aya Expanse 32b has downloadable open weights; the other is API-only.

Side by side

C4ai Aya Expanse 32b and Gemini 3.5 Flash specifications
C4ai Aya Expanse 32bGemini 3.5 Flash
ProviderCohereGoogle
Noometry Index35.954.2
Released2024-10-242026-05-19
WeightsOpenProprietary
Context window128K1.05M
Max output4K66K
Input $ / M tokens—$1.50
Output $ / M tokens—$9
Results tracked1854

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 3.5 Flash leads

C4ai Aya Expanse 32b: 34.8 (#231), Gemini 3.5 Flash: 49.4 (#49)

Coding benchmarks
BenchmarkC4ai Aya Expanse 32bGemini 3.5 Flash
LMArena Coding11971492
SWE-bench Verified—79.3%
DeepSWE—37.4%
LMArena WebDev—1499
SciCode—53.1%
WeirdML—62.6%
ALE-Bench—911.02

Agentic & Tool Use Not comparable

C4ai Aya Expanse 32b: —, Gemini 3.5 Flash: 24.7 (#114)

Agentic & Tool Use benchmarks
BenchmarkC4ai Aya Expanse 32bGemini 3.5 Flash
APEX-Agents—27.5%
GBAEval—6.7%
GDP.pdf—14%
Vending-Bench 2—5,396

Reasoning Gemini 3.5 Flash leads

C4ai Aya Expanse 32b: 23.3 (#180), Gemini 3.5 Flash: 62.8 (#18)

Reasoning benchmarks
BenchmarkC4ai Aya Expanse 32bGemini 3.5 Flash
LMArena Hard Prompts11931488
ARC-AGI-2—72.1%
SimpleBench—76.7%
NYT Connections (extended)—92.6%
ARC-AGI-1—92.5%
CritPt—13.1%
Chess Puzzles—50%
EnigmaEval—25.4%
EBR-Bench—4.8%
Mystery Game Puzzles—32%
DTBench—94.7%
LMCA—47.1%
Surface Evolver Bench—58.1%
Epoch Capabilities Index—154.46
ForecastBench—59

Math Gemini 3.5 Flash leads

C4ai Aya Expanse 32b: 34.0 (#197), Gemini 3.5 Flash: 60.7 (#36)

Math benchmarks
BenchmarkC4ai Aya Expanse 32bGemini 3.5 Flash
LMArena Math12001504
FrontierMath (Tiers 1-3)—62.8%
FrontierMath Tier 4—26.8%
MathArena Final-Answer Competitions—76.3%
OTIS Mock AIME 2024-2025—95.6%
ProofBench—31%
FrontierMath (Feb 2025 set)—39%
FrontierMath Tier 4 (v1)—14.6%

Knowledge Gemini 3.5 Flash leads

C4ai Aya Expanse 32b: 33.2 (#206), Gemini 3.5 Flash: 66.3 (#11)

Knowledge benchmarks
BenchmarkC4ai Aya Expanse 32bGemini 3.5 Flash
LMArena Expert11821495
GPQA Diamond—92.8%
SimpleQA Verified—66.2%
Vectara Hallucination Rate10.9%—

Multimodal Not comparable

C4ai Aya Expanse 32b: —, Gemini 3.5 Flash: 45.7 (#15)

Multimodal benchmarks
BenchmarkC4ai Aya Expanse 32bGemini 3.5 Flash
LMArena Vision—1310
Blueprint-Bench 2—33.6%
LMArena Document—1463

Multilingual Gemini 3.5 Flash leads

C4ai Aya Expanse 32b: 38.4 (#230), Gemini 3.5 Flash: 57.0 (#13)

Multilingual benchmarks
BenchmarkC4ai Aya Expanse 32bGemini 3.5 Flash
LMArena Non-English12131476
LMArena Chinese12111526
LMArena French12481490
LMArena German11991492
LMArena Japanese11631486
LMArena Korean11581451
LMArena Russian12271493
LMArena Spanish11931480

Instruction Following Gemini 3.5 Flash leads

C4ai Aya Expanse 32b: 62.6 (#237), Gemini 3.5 Flash: 77.0 (#30)

Instruction Following benchmarks
BenchmarkC4ai Aya Expanse 32bGemini 3.5 Flash
LMArena Instruction Following11961467

Long Context Gemini 3.5 Flash leads

C4ai Aya Expanse 32b: 37.2 (#220), Gemini 3.5 Flash: 45.4 (#38)

Long Context benchmarks
BenchmarkC4ai Aya Expanse 32bGemini 3.5 Flash
LMArena Longer Query12281482

Writing & Preference Gemini 3.5 Flash leads

C4ai Aya Expanse 32b: 42.2 (#235), Gemini 3.5 Flash: 65.5 (#47)

Writing & Preference benchmarks
BenchmarkC4ai Aya Expanse 32bGemini 3.5 Flash
LMArena Text12241482
LMArena Creative Writing12001470
LMArena Multi-Turn11901481
EQ-Bench 4—1087

Frequently asked questions

Is C4ai Aya Expanse 32b better than Gemini 3.5 Flash?

Gemini 3.5 Flash is the stronger model overall, scoring 54.2 to 35.9 on the Noometry Index.

Is C4ai Aya Expanse 32b or Gemini 3.5 Flash better for coding?

Gemini 3.5 Flash scores higher on coding benchmarks: 49.4 versus 34.8 in the Noometry coding category.

Which has the bigger context window?

Gemini 3.5 Flash does, with 1.05M tokens against 128K.

How many benchmarks do C4ai Aya Expanse 32b and Gemini 3.5 Flash share?

17 benchmarks have published results for both models. C4ai Aya Expanse 32b has 18 scored results on Noometry and Gemini 3.5 Flash has 54.

Related comparisons

Go deeper