Model comparison

Command A vs Hy3

Hy3 is the stronger model overall, scoring 44.2 to 36.5 on the Noometry Index.

Last verified . 17 shared benchmarks.

Command A Cohere

36.5

Rank #215 Confirmed

Hy3 Tencent

44.2

Rank #79 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Command A scores higher in 0 categories and Hy3 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Hy3 leads 46.8 to 27.2.
  • Hy3 is cheaper at $0.0825 / $0.33 per million input/output tokens, against $2.50 / $10 for Command A.
  • Hy3 accepts more context: 262K tokens versus 256K.

Side by side

Command A and Hy3 specifications
Command AHy3
ProviderCohereTencent
Noometry Index36.544.2
Released2025-03-132026-07-06
WeightsOpenOpen
Context window256K262K
Max output8K128K
Input $ / M tokens$2.50$0.0825
Output $ / M tokens$10$0.33
Results tracked2419

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hy3 leads

Command A: 27.2 (#322), Hy3: 46.8 (#63)

Coding benchmarks
BenchmarkCommand AHy3
LMArena Coding13301464
Aider Polyglot12%—
LMArena WebDev—1508

Agentic & Tool Use Not comparable

Command A: 35.9 (#40), Hy3: —

Agentic & Tool Use benchmarks
BenchmarkCommand AHy3
Berkeley Function Calling Leaderboard57.1%—

Reasoning Hy3 leads

Command A: 18.3 (#283), Hy3: 26.1 (#136)

Reasoning benchmarks
BenchmarkCommand AHy3
LMArena Hard Prompts13261447
Kagi LLM Benchmark28.8%—
NYT Connections (extended)—41.2%
DTBench61.3%—
LMCA10.3%—

Math Hy3 leads

Command A: 36.2 (#171), Hy3: 40.1 (#93)

Math benchmarks
BenchmarkCommand AHy3
LMArena Math13001475

Knowledge Hy3 leads

Command A: 37.1 (#159), Hy3: 40.8 (#114)

Knowledge benchmarks
BenchmarkCommand AHy3
LMArena Expert12951460
Vectara Hallucination Rate9.3%—

Multilingual Hy3 leads

Command A: 45.3 (#170), Hy3: 53.5 (#65)

Multilingual benchmarks
BenchmarkCommand AHy3
LMArena Non-English13131426
LMArena Chinese13271493
LMArena French13511461
LMArena German13411439
LMArena Japanese12851392
LMArena Korean12851395
LMArena Russian13141432
LMArena Spanish13471456

Instruction Following Hy3 leads

Command A: 69.1 (#177), Hy3: 75.1 (#70)

Instruction Following benchmarks
BenchmarkCommand AHy3
LMArena Instruction Following13091426

Long Context Hy3 leads

Command A: 40.6 (#151), Hy3: 44.1 (#75)

Long Context benchmarks
BenchmarkCommand AHy3
LMArena Longer Query13341442

Writing & Preference Hy3 leads

Command A: 47.6 (#208), Hy3: 62.2 (#81)

Writing & Preference benchmarks
BenchmarkCommand AHy3
LMArena Text13311439
LMArena Creative Writing13191402
LMArena Multi-Turn13391436
EQ-Bench Creative Writing1145—

Frequently asked questions

Is Command A better than Hy3?

Hy3 is the stronger model overall, scoring 44.2 to 36.5 on the Noometry Index.

Which is cheaper, Command A or Hy3?

Hy3 is cheaper. It lists at $0.0825 per million input tokens and $0.33 per million output tokens; Command A lists at $2.50 and $10.

Is Command A or Hy3 better for coding?

Hy3 scores higher on coding benchmarks: 46.8 versus 27.2 in the Noometry coding category.

Which has the bigger context window?

Hy3 does, with 262K tokens against 256K.

How many benchmarks do Command A and Hy3 share?

17 benchmarks have published results for both models. Command A has 24 scored results on Noometry and Hy3 has 19.

Related comparisons

Go deeper