Model comparison

Command R vs Inkling

Inkling is the stronger model overall, scoring 44.1 to 31.4 on the Noometry Index. Command R costs 9.8× less per token, which makes it the better buy when Inkling's lead doesn't matter for your workload.

Last verified . 19 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Inkling Thinking Machines Lab

44.1

Rank #80 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Command R scores higher in 0 categories and Inkling in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Inkling leads 65.2 to 38.2.
  • The biggest single-benchmark swing is DTBench: 46.4% for Command R and 87.5% for Inkling.
  • Command R is cheaper at $0.15 / $0.60 per million input/output tokens, against $1.87 / $4.68 for Inkling.
  • Command R accepts more context: 128K tokens versus 66K.

Side by side

Command R and Inkling specifications
Command RInkling
ProviderCohereThinking Machines Lab
Noometry Index31.444.1
Released2024-08-302026-07-15
WeightsOpenOpen
Context window128K66K
Max output4K66K
Input $ / M tokens$0.15$1.87
Output $ / M tokens$0.60$4.68
Results tracked2941

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Inkling leads

Command R: 29.3 (#306), Inkling: 34.5 (#234)

Coding benchmarks
BenchmarkCommand RInkling
LMArena Coding11691464
FrontierCode—14%
LMArena WebDev—1413
FrontierSWE—4.1%
SciCode—47%
WeirdML—32.3%
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
BigCodeBench Complete45.2%—
ALE-Bench—946

Agentic & Tool Use Not comparable

Command R: —, Inkling: 29.6 (#85)

Agentic & Tool Use benchmarks
BenchmarkCommand RInkling
APEX-Agents—33.8%
τ²-bench Banking—25%

Reasoning Inkling leads

Command R: 13.8 (#331), Inkling: 40.4 (#56)

Reasoning benchmarks
BenchmarkCommand RInkling
LMArena Hard Prompts11641451
DTBench46.4%87.5%
LMCA9.2%37.6%
ARC-AGI-2—36.5%
SimpleBench—50%
ARC-AGI-1—79.5%
CritPt—5.4%
Chess Puzzles—21%
LiveBench Reasoning21.9%—
LiveBench Data Analysis33.3%—
Epoch Capabilities Index—148.54
LiveBench27.5%—

Math Inkling leads

Command R: 28.0 (#246), Inkling: 31.3 (#225)

Math benchmarks
BenchmarkCommand RInkling
LMArena Math11551479
FrontierMath (Tiers 1-3)—33.3%
FrontierMath Tier 4—4.9%
OTIS Mock AIME 2024-2025—88.9%
ProofBench—0%
LiveBench Math19.4%—

Knowledge Inkling leads

Command R: 31.0 (#221), Inkling: 55.1 (#49)

Knowledge benchmarks
BenchmarkCommand RInkling
LMArena Expert11381465
GPQA Diamond—88.3%
SimpleQA Verified—40.3%
MMLU65.2%—

Multilingual Inkling leads

Command R: 35.7 (#245), Inkling: 54.0 (#52)

Multilingual benchmarks
BenchmarkCommand RInkling
LMArena Non-English11741434
LMArena Chinese11821490
LMArena French11621458
LMArena German11761446
LMArena Japanese11431429
LMArena Korean11631404
LMArena Russian11741429
LMArena Spanish11511448

Instruction Following Inkling leads

Command R: 58.1 (#261), Inkling: 75.1 (#71)

Instruction Following benchmarks
BenchmarkCommand RInkling
LMArena Instruction Following11671426
LiveBench Instruction Following55.6%—

Long Context Inkling leads

Command R: 36.3 (#231), Inkling: 43.8 (#86)

Long Context benchmarks
BenchmarkCommand RInkling
LMArena Longer Query11981434

Writing & Preference Inkling leads

Command R: 38.2 (#254), Inkling: 65.2 (#51)

Writing & Preference benchmarks
BenchmarkCommand RInkling
LMArena Text11871441
LMArena Creative Writing11701387
LMArena Multi-Turn11631436
EQ-Bench Creative Writing—1611
EQ-Bench 4—1226
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Inkling?

Inkling is the stronger model overall, scoring 44.1 to 31.4 on the Noometry Index. Command R costs 9.8× less per token, which makes it the better buy when Inkling's lead doesn't matter for your workload.

Which is cheaper, Command R or Inkling?

Command R is cheaper. It lists at $0.15 per million input tokens and $0.60 per million output tokens; Inkling lists at $1.87 and $4.68.

Is Command R or Inkling better for coding?

Inkling scores higher on coding benchmarks: 34.5 versus 29.3 in the Noometry coding category.

Which has the bigger context window?

Command R does, with 128K tokens against 66K.

How many benchmarks do Command R and Inkling share?

19 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Inkling has 41.

Related comparisons

Go deeper