Model comparison

Command R+ vs Nvidia Llama 3.3 Nemotron Super 49b v1.5

Nvidia Llama 3.3 Nemotron Super 49b v1.5 is the stronger model overall, scoring 40.3 to 32.4 on the Noometry Index.

Last verified . 12 shared benchmarks.

Command R+ Cohere

32.4

Rank #257 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Command R+ scores higher in 0 categories and Nvidia Llama 3.3 Nemotron Super 49b v1.5 in 8 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads 26.8 to 9.2.
  • Nvidia Llama 3.3 Nemotron Super 49b v1.5 is cheaper at $0.40 / $0.40 per million input/output tokens, against $2.50 / $10 for Command R+.
  • Nvidia Llama 3.3 Nemotron Super 49b v1.5 accepts more context: 131K tokens versus 128K.

Side by side

Command R+ and Nvidia Llama 3.3 Nemotron Super 49b v1.5 specifications
Command R+Nvidia Llama 3.3 Nemotron Super 49b v1.5
ProviderCohereNVIDIA
Noometry Index32.440.3
Released2024-08-302025-07-25
WeightsOpenOpen
Context window128K131K
Max output4K131K
Input $ / M tokens$2.50$0.40
Output $ / M tokens$10$0.40
Results tracked3412

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Command R+: 29.1 (#309), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 39.8 (#154)

Coding benchmarks
BenchmarkCommand R+Nvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Coding11871355
BigCodeBench Instruct33.8%—
LiveBench Coding19.1%—
BigCodeBench Complete41.9%—
HumanEval+56.7%—
MBPP+63.5%—

Reasoning Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Command R+: 9.2 (#344), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 26.8 (#128)

Reasoning benchmarks
BenchmarkCommand R+Nvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Hard Prompts11861336
SimpleBench17.4%—
LiveBench Reasoning24.8%—
DTBench54.9%—
LiveBench Data Analysis38.1%—
LMCA5%—
Epoch Capabilities Index119.34—
LiveBench31.8%—

Math Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Command R+: 28.9 (#242), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 38.2 (#141)

Math benchmarks
BenchmarkCommand R+Nvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Math11881392
LiveBench Math21.3%—

Knowledge Too close to call

Command R+: 36.4 (#169), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 36.7 (#165)

Knowledge benchmarks
BenchmarkCommand R+Nvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Expert11741330
Vectara Hallucination Rate6.9%—
MMLU69.4%—

Multilingual Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Command R+: 38.6 (#227), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 45.5 (#168)

Multilingual benchmarks
BenchmarkCommand R+Nvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Non-English12161316
LMArena Japanese11661300
LMArena Russian12271332
LMArena Chinese1226—
LMArena French1209—
LMArena German1216—
LMArena Korean1138—
LMArena Spanish1189—

Instruction Following Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Command R+: 60.0 (#254), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 68.6 (#188)

Instruction Following benchmarks
BenchmarkCommand R+Nvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Instruction Following11971299
LiveBench Instruction Following57.6%—

Long Context Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Command R+: 37.3 (#219), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 40.0 (#164)

Long Context benchmarks
BenchmarkCommand R+Nvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Longer Query12301315

Writing & Preference Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Command R+: 43.5 (#228), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 53.1 (#159)

Writing & Preference benchmarks
BenchmarkCommand R+Nvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Text12291338
LMArena Creative Writing12351307
LMArena Multi-Turn12131334
LiveBench Language29.7%—

Frequently asked questions

Is Command R+ better than Nvidia Llama 3.3 Nemotron Super 49b v1.5?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 is the stronger model overall, scoring 40.3 to 32.4 on the Noometry Index.

Which is cheaper, Command R+ or Nvidia Llama 3.3 Nemotron Super 49b v1.5?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 is cheaper. It lists at $0.40 per million input tokens and $0.40 per million output tokens; Command R+ lists at $2.50 and $10.

Is Command R+ or Nvidia Llama 3.3 Nemotron Super 49b v1.5 better for coding?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 scores higher on coding benchmarks: 39.8 versus 29.1 in the Noometry coding category.

Which has the bigger context window?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 does, with 131K tokens against 128K.

How many benchmarks do Command R+ and Nvidia Llama 3.3 Nemotron Super 49b v1.5 share?

12 benchmarks have published results for both models. Command R+ has 34 scored results on Noometry and Nvidia Llama 3.3 Nemotron Super 49b v1.5 has 12.

Related comparisons

Go deeper