Model comparison

C4ai Aya Expanse 8b vs Inkling

Inkling is the stronger model overall, scoring 44.1 to 34.9 on the Noometry Index.

Last verified . 14 shared benchmarks.

C4ai Aya Expanse 8b Cohere

34.9

Rank #229 Confirmed

Inkling Thinking Machines Lab

44.1

Rank #80 Confirmed

Summary

  • They share 14 benchmarks with published results for both. C4ai Aya Expanse 8b scores higher in 1 category and Inkling in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Inkling leads 65.2 to 39.0.

Side by side

C4ai Aya Expanse 8b and Inkling specifications
C4ai Aya Expanse 8bInkling
ProviderCohereThinking Machines Lab
Noometry Index34.944.1
Released—2026-07-15
WeightsOpenOpen
Context window—66K
Max output—66K
Input $ / M tokens—$1.87
Output $ / M tokens—$4.68
Results tracked1541

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

C4ai Aya Expanse 8b: 33.8 (#252), Inkling: 34.5 (#234)

Coding benchmarks
BenchmarkC4ai Aya Expanse 8bInkling
LMArena Coding11601464
FrontierCode—14%
LMArena WebDev—1413
FrontierSWE—4.1%
SciCode—47%
WeirdML—32.3%
ALE-Bench—946

Agentic & Tool Use Not comparable

C4ai Aya Expanse 8b: —, Inkling: 29.6 (#85)

Agentic & Tool Use benchmarks
BenchmarkC4ai Aya Expanse 8bInkling
APEX-Agents—33.8%
τ²-bench Banking—25%

Reasoning Inkling leads

C4ai Aya Expanse 8b: 22.4 (#197), Inkling: 40.4 (#56)

Reasoning benchmarks
BenchmarkC4ai Aya Expanse 8bInkling
LMArena Hard Prompts11561451
ARC-AGI-2—36.5%
SimpleBench—50%
ARC-AGI-1—79.5%
CritPt—5.4%
Chess Puzzles—21%
DTBench—87.5%
LMCA—37.6%
Epoch Capabilities Index—148.54

Math C4ai Aya Expanse 8b leads

C4ai Aya Expanse 8b: 33.3 (#203), Inkling: 31.3 (#225)

Math benchmarks
BenchmarkC4ai Aya Expanse 8bInkling
LMArena Math11681479
FrontierMath (Tiers 1-3)—33.3%
FrontierMath Tier 4—4.9%
OTIS Mock AIME 2024-2025—88.9%
ProofBench—0%

Knowledge Inkling leads

C4ai Aya Expanse 8b: 33.6 (#201), Inkling: 55.1 (#49)

Knowledge benchmarks
BenchmarkC4ai Aya Expanse 8bInkling
LMArena Expert11531465
GPQA Diamond—88.3%
SimpleQA Verified—40.3%
Vectara Hallucination Rate9.5%—

Multilingual Inkling leads

C4ai Aya Expanse 8b: 36.1 (#242), Inkling: 54.0 (#52)

Multilingual benchmarks
BenchmarkC4ai Aya Expanse 8bInkling
LMArena Non-English11801434
LMArena Chinese11811490
LMArena German11881446
LMArena Japanese11201429
LMArena Russian11971429
LMArena French—1458
LMArena Korean—1404
LMArena Spanish—1448

Instruction Following Inkling leads

C4ai Aya Expanse 8b: 60.1 (#253), Inkling: 75.1 (#71)

Instruction Following benchmarks
BenchmarkC4ai Aya Expanse 8bInkling
LMArena Instruction Following11551426

Long Context Inkling leads

C4ai Aya Expanse 8b: 36.1 (#236), Inkling: 43.8 (#86)

Long Context benchmarks
BenchmarkC4ai Aya Expanse 8bInkling
LMArena Longer Query11911434

Writing & Preference Inkling leads

C4ai Aya Expanse 8b: 39.0 (#250), Inkling: 65.2 (#51)

Writing & Preference benchmarks
BenchmarkC4ai Aya Expanse 8bInkling
LMArena Text11851441
LMArena Creative Writing11681387
LMArena Multi-Turn11601436
EQ-Bench Creative Writing—1611
EQ-Bench 4—1226

Frequently asked questions

Is C4ai Aya Expanse 8b better than Inkling?

Inkling is the stronger model overall, scoring 44.1 to 34.9 on the Noometry Index.

Is C4ai Aya Expanse 8b or Inkling better for coding?

They score almost the same on coding (33.8 vs 34.5); test both on your own repository before choosing.

How many benchmarks do C4ai Aya Expanse 8b and Inkling share?

14 benchmarks have published results for both models. C4ai Aya Expanse 8b has 15 scored results on Noometry and Inkling has 41.

Related comparisons

Go deeper