Model comparison

C4ai Aya Expanse 8b vs Trinity Large Thinking

Trinity Large Thinking is the stronger model overall, scoring 38.6 to 34.9 on the Noometry Index.

Last verified . 15 shared benchmarks.

C4ai Aya Expanse 8b Cohere

34.9

Rank #229 Confirmed

Trinity Large Thinking Arcee AI

38.6

Rank #185 Confirmed

Summary

  • They share 15 benchmarks with published results for both. C4ai Aya Expanse 8b scores higher in 1 category and Trinity Large Thinking in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Trinity Large Thinking leads 53.8 to 39.0.

Side by side

C4ai Aya Expanse 8b and Trinity Large Thinking specifications
C4ai Aya Expanse 8bTrinity Large Thinking
ProviderCohereArcee AI
Noometry Index34.938.6
Released—2026-04-01
WeightsOpenOpen
Context window—262K
Max output—80K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.80
Results tracked1524

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

C4ai Aya Expanse 8b: 33.8 (#252), Trinity Large Thinking: 34.1 (#244)

Coding benchmarks
BenchmarkC4ai Aya Expanse 8bTrinity Large Thinking
LMArena Coding11601381
LMArena WebDev—1238
SciCode—36.1%

Reasoning C4ai Aya Expanse 8b leads

C4ai Aya Expanse 8b: 22.4 (#197), Trinity Large Thinking: 16.9 (#298)

Reasoning benchmarks
BenchmarkC4ai Aya Expanse 8bTrinity Large Thinking
LMArena Hard Prompts11561350
NYT Connections (extended)—16.5%
CritPt—0.9%
Thematic Generalization—41.6%
Surface Evolver Bench—15.6%

Math Trinity Large Thinking leads

C4ai Aya Expanse 8b: 33.3 (#203), Trinity Large Thinking: 37.6 (#149)

Math benchmarks
BenchmarkC4ai Aya Expanse 8bTrinity Large Thinking
LMArena Math11681366

Knowledge Trinity Large Thinking leads

C4ai Aya Expanse 8b: 33.6 (#201), Trinity Large Thinking: 40.9 (#113)

Knowledge benchmarks
BenchmarkC4ai Aya Expanse 8bTrinity Large Thinking
Vectara Hallucination Rate9.5%6.9%
LMArena Expert11531360

Multilingual Trinity Large Thinking leads

C4ai Aya Expanse 8b: 36.1 (#242), Trinity Large Thinking: 46.2 (#160)

Multilingual benchmarks
BenchmarkC4ai Aya Expanse 8bTrinity Large Thinking
LMArena Non-English11801325
LMArena Chinese11811373
LMArena German11881356
LMArena Japanese11201311
LMArena Russian11971337
LMArena French—1374
LMArena Korean—1306
LMArena Spanish—1357

Instruction Following Trinity Large Thinking leads

C4ai Aya Expanse 8b: 60.1 (#253), Trinity Large Thinking: 70.5 (#162)

Instruction Following benchmarks
BenchmarkC4ai Aya Expanse 8bTrinity Large Thinking
LMArena Instruction Following11551334

Long Context Trinity Large Thinking leads

C4ai Aya Expanse 8b: 36.1 (#236), Trinity Large Thinking: 41.3 (#144)

Long Context benchmarks
BenchmarkC4ai Aya Expanse 8bTrinity Large Thinking
LMArena Longer Query11911355

Writing & Preference Trinity Large Thinking leads

C4ai Aya Expanse 8b: 39.0 (#250), Trinity Large Thinking: 53.8 (#158)

Writing & Preference benchmarks
BenchmarkC4ai Aya Expanse 8bTrinity Large Thinking
LMArena Text11851340
LMArena Creative Writing11681320
LMArena Multi-Turn11601342

Frequently asked questions

Is C4ai Aya Expanse 8b better than Trinity Large Thinking?

Trinity Large Thinking is the stronger model overall, scoring 38.6 to 34.9 on the Noometry Index.

Is C4ai Aya Expanse 8b or Trinity Large Thinking better for coding?

They score almost the same on coding (33.8 vs 34.1); test both on your own repository before choosing.

How many benchmarks do C4ai Aya Expanse 8b and Trinity Large Thinking share?

15 benchmarks have published results for both models. C4ai Aya Expanse 8b has 15 scored results on Noometry and Trinity Large Thinking has 24.

Related comparisons

Go deeper