Model comparison

Command A vs Llama 3.3 Nemotron 49b Super v1

Llama 3.3 Nemotron 49b Super v1 is the stronger model overall, scoring 40.1 to 36.5 on the Noometry Index.

Last verified . 10 shared benchmarks.

Command A Cohere

36.5

Rank #215 Confirmed

Summary

  • They share 10 benchmarks with published results for both. Command A scores higher in 3 categories and Llama 3.3 Nemotron 49b Super v1 in 3 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Llama 3.3 Nemotron 49b Super v1 leads 37.9 to 27.2.

Side by side

Command A and Llama 3.3 Nemotron 49b Super v1 specifications
Command ALlama 3.3 Nemotron 49b Super v1
ProviderCohereNVIDIA
Noometry Index36.540.1
Released2025-03-13—
WeightsOpenOpen
Context window256K—
Max output8K—
Input $ / M tokens$2.50—
Output $ / M tokens$10—
Results tracked2410

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Llama 3.3 Nemotron 49b Super v1 leads

Command A: 27.2 (#322), Llama 3.3 Nemotron 49b Super v1: 37.9 (#186)

Coding benchmarks
BenchmarkCommand ALlama 3.3 Nemotron 49b Super v1
LMArena Coding13301296
Aider Polyglot12%—

Agentic & Tool Use Not comparable

Command A: 35.9 (#40), Llama 3.3 Nemotron 49b Super v1: —

Agentic & Tool Use benchmarks
BenchmarkCommand ALlama 3.3 Nemotron 49b Super v1
Berkeley Function Calling Leaderboard57.1%—

Reasoning Llama 3.3 Nemotron 49b Super v1 leads

Command A: 18.3 (#283), Llama 3.3 Nemotron 49b Super v1: 26.2 (#135)

Reasoning benchmarks
BenchmarkCommand ALlama 3.3 Nemotron 49b Super v1
LMArena Hard Prompts13261311
Kagi LLM Benchmark28.8%—
DTBench61.3%—
LMCA10.3%—

Math Not comparable

Command A: 36.2 (#171), Llama 3.3 Nemotron 49b Super v1: —

Math benchmarks
BenchmarkCommand ALlama 3.3 Nemotron 49b Super v1
LMArena Math1300—

Knowledge Not comparable

Command A: 37.1 (#159), Llama 3.3 Nemotron 49b Super v1: —

Knowledge benchmarks
BenchmarkCommand ALlama 3.3 Nemotron 49b Super v1
Vectara Hallucination Rate9.3%—
LMArena Expert1295—

Multilingual Command A leads

Command A: 45.3 (#170), Llama 3.3 Nemotron 49b Super v1: 41.1 (#211)

Multilingual benchmarks
BenchmarkCommand ALlama 3.3 Nemotron 49b Super v1
LMArena Non-English13131253
LMArena Chinese13271277
LMArena Russian13141269
LMArena French1351—
LMArena German1341—
LMArena Japanese1285—
LMArena Korean1285—
LMArena Spanish1347—

Instruction Following Too close to call

Command A: 69.1 (#177), Llama 3.3 Nemotron 49b Super v1: 68.3 (#189)

Instruction Following benchmarks
BenchmarkCommand ALlama 3.3 Nemotron 49b Super v1
LMArena Instruction Following13091293

Long Context Command A leads

Command A: 40.6 (#151), Llama 3.3 Nemotron 49b Super v1: 39.5 (#176)

Long Context benchmarks
BenchmarkCommand ALlama 3.3 Nemotron 49b Super v1
LMArena Longer Query13341299

Writing & Preference Llama 3.3 Nemotron 49b Super v1 leads

Command A: 47.6 (#208), Llama 3.3 Nemotron 49b Super v1: 50.8 (#179)

Writing & Preference benchmarks
BenchmarkCommand ALlama 3.3 Nemotron 49b Super v1
LMArena Text13311308
LMArena Creative Writing13191288
LMArena Multi-Turn13391315
EQ-Bench Creative Writing1145—

Frequently asked questions

Is Command A better than Llama 3.3 Nemotron 49b Super v1?

Llama 3.3 Nemotron 49b Super v1 is the stronger model overall, scoring 40.1 to 36.5 on the Noometry Index.

Is Command A or Llama 3.3 Nemotron 49b Super v1 better for coding?

Llama 3.3 Nemotron 49b Super v1 scores higher on coding benchmarks: 37.9 versus 27.2 in the Noometry coding category.

How many benchmarks do Command A and Llama 3.3 Nemotron 49b Super v1 share?

10 benchmarks have published results for both models. Command A has 24 scored results on Noometry and Llama 3.3 Nemotron 49b Super v1 has 10.

Related comparisons

Go deeper