Model comparison

Command A vs Step 1o Turbo 202506

Step 1o Turbo 202506 is the stronger model overall, scoring 39.7 to 36.5 on the Noometry Index.

Last verified . 13 shared benchmarks.

Command A Cohere

36.5

Rank #215 Confirmed

Step 1o Turbo 202506 StepFun

39.7

Rank #160 Confirmed

Summary

  • They share 13 benchmarks with published results for both. Command A scores higher in 2 categories and Step 1o Turbo 202506 in 6 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Step 1o Turbo 202506 leads 39.2 to 27.2.
  • Command A has downloadable open weights; the other is API-only.

Side by side

Command A and Step 1o Turbo 202506 specifications
Command AStep 1o Turbo 202506
ProviderCohereStepFun
Noometry Index36.539.7
Released2025-03-13—
WeightsOpenProprietary
Context window256K—
Max output8K—
Input $ / M tokens$2.50—
Output $ / M tokens$10—
Results tracked2414

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Step 1o Turbo 202506 leads

Command A: 27.2 (#322), Step 1o Turbo 202506: 39.2 (#160)

Coding benchmarks
BenchmarkCommand AStep 1o Turbo 202506
LMArena Coding13301339
Aider Polyglot12%—

Agentic & Tool Use Not comparable

Command A: 35.9 (#40), Step 1o Turbo 202506: —

Agentic & Tool Use benchmarks
BenchmarkCommand AStep 1o Turbo 202506
Berkeley Function Calling Leaderboard57.1%—

Reasoning Step 1o Turbo 202506 leads

Command A: 18.3 (#283), Step 1o Turbo 202506: 26.8 (#129)

Reasoning benchmarks
BenchmarkCommand AStep 1o Turbo 202506
LMArena Hard Prompts13261335
Kagi LLM Benchmark28.8%—
DTBench61.3%—
LMCA10.3%—

Math Too close to call

Command A: 36.2 (#171), Step 1o Turbo 202506: 36.6 (#164)

Math benchmarks
BenchmarkCommand AStep 1o Turbo 202506
LMArena Math13001318

Knowledge Command A leads

Command A: 37.1 (#159), Step 1o Turbo 202506: 36.1 (#176)

Knowledge benchmarks
BenchmarkCommand AStep 1o Turbo 202506
LMArena Expert12951308
Vectara Hallucination Rate9.3%—

Multimodal Not comparable

Command A: —, Step 1o Turbo 202506: 36.1 (#80)

Multimodal benchmarks
BenchmarkCommand AStep 1o Turbo 202506
LMArena Vision—1186

Multilingual Too close to call

Command A: 45.3 (#170), Step 1o Turbo 202506: 45.3 (#173)

Multilingual benchmarks
BenchmarkCommand AStep 1o Turbo 202506
LMArena Non-English13131313
LMArena Chinese13271380
LMArena German13411314
LMArena Russian13141327
LMArena French1351—
LMArena Japanese1285—
LMArena Korean1285—
LMArena Spanish1347—

Instruction Following Too close to call

Command A: 69.1 (#177), Step 1o Turbo 202506: 69.2 (#176)

Instruction Following benchmarks
BenchmarkCommand AStep 1o Turbo 202506
LMArena Instruction Following13091310

Long Context Too close to call

Command A: 40.6 (#151), Step 1o Turbo 202506: 41.0 (#146)

Long Context benchmarks
BenchmarkCommand AStep 1o Turbo 202506
LMArena Longer Query13341348

Writing & Preference Step 1o Turbo 202506 leads

Command A: 47.6 (#208), Step 1o Turbo 202506: 53.1 (#160)

Writing & Preference benchmarks
BenchmarkCommand AStep 1o Turbo 202506
LMArena Text13311336
LMArena Creative Writing13191307
LMArena Multi-Turn13391340
EQ-Bench Creative Writing1145—

Frequently asked questions

Is Command A better than Step 1o Turbo 202506?

Step 1o Turbo 202506 is the stronger model overall, scoring 39.7 to 36.5 on the Noometry Index.

Is Command A or Step 1o Turbo 202506 better for coding?

Step 1o Turbo 202506 scores higher on coding benchmarks: 39.2 versus 27.2 in the Noometry coding category.

How many benchmarks do Command A and Step 1o Turbo 202506 share?

13 benchmarks have published results for both models. Command A has 24 scored results on Noometry and Step 1o Turbo 202506 has 14.

Related comparisons

Go deeper