Model comparison

Command R vs Phi 3 Mini 4k Instruct June 2024

Command R and Phi 3 Mini 4k Instruct June 2024 score almost the same on the Noometry Index (31.4 vs 31.3), so choose on price, context window or the category you care about most.

Last verified . 15 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Summary

  • They share 15 benchmarks with published results for both. Command R scores higher in 5 categories and Phi 3 Mini 4k Instruct June 2024 in 3 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in multilingual, where Command R leads 35.7 to 25.9.

Side by side

Command R and Phi 3 Mini 4k Instruct June 2024 specifications
Command RPhi 3 Mini 4k Instruct June 2024
ProviderCohereMicrosoft
Noometry Index31.431.3
Released2024-08-30—
WeightsOpenOpen
Context window128K—
Max output4K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.60—
Results tracked2915

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Phi 3 Mini 4k Instruct June 2024 leads

Command R: 29.3 (#306), Phi 3 Mini 4k Instruct June 2024: 31.8 (#279)

Coding benchmarks
BenchmarkCommand RPhi 3 Mini 4k Instruct June 2024
LMArena Coding11691093
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
BigCodeBench Complete45.2%—

Reasoning Phi 3 Mini 4k Instruct June 2024 leads

Command R: 13.8 (#331), Phi 3 Mini 4k Instruct June 2024: 20.8 (#231)

Reasoning benchmarks
BenchmarkCommand RPhi 3 Mini 4k Instruct June 2024
LMArena Hard Prompts11641087
LiveBench Reasoning21.9%—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
LiveBench27.5%—

Math Phi 3 Mini 4k Instruct June 2024 leads

Command R: 28.0 (#246), Phi 3 Mini 4k Instruct June 2024: 33.0 (#208)

Math benchmarks
BenchmarkCommand RPhi 3 Mini 4k Instruct June 2024
LMArena Math11551152
LiveBench Math19.4%—

Knowledge Command R leads

Command R: 31.0 (#221), Phi 3 Mini 4k Instruct June 2024: 28.6 (#244)

Knowledge benchmarks
BenchmarkCommand RPhi 3 Mini 4k Instruct June 2024
LMArena Expert11381051
MMLU65.2%—

Multilingual Command R leads

Command R: 35.7 (#245), Phi 3 Mini 4k Instruct June 2024: 25.9 (#282)

Multilingual benchmarks
BenchmarkCommand RPhi 3 Mini 4k Instruct June 2024
LMArena Non-English11741013
LMArena Chinese11821033
LMArena German11761031
LMArena Japanese1143954
LMArena Korean1163880
LMArena Russian11741019
LMArena French1162—
LMArena Spanish1151—

Instruction Following Command R leads

Command R: 58.1 (#261), Phi 3 Mini 4k Instruct June 2024: 54.1 (#282)

Instruction Following benchmarks
BenchmarkCommand RPhi 3 Mini 4k Instruct June 2024
LMArena Instruction Following11671058
LiveBench Instruction Following55.6%—

Long Context Command R leads

Command R: 36.3 (#231), Phi 3 Mini 4k Instruct June 2024: 31.7 (#277)

Long Context benchmarks
BenchmarkCommand RPhi 3 Mini 4k Instruct June 2024
LMArena Longer Query11981042

Writing & Preference Command R leads

Command R: 38.2 (#254), Phi 3 Mini 4k Instruct June 2024: 29.6 (#294)

Writing & Preference benchmarks
BenchmarkCommand RPhi 3 Mini 4k Instruct June 2024
LMArena Text11871080
LMArena Creative Writing11701045
LMArena Multi-Turn11631049
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Phi 3 Mini 4k Instruct June 2024?

Command R and Phi 3 Mini 4k Instruct June 2024 score almost the same on the Noometry Index (31.4 vs 31.3), so choose on price, context window or the category you care about most.

Is Command R or Phi 3 Mini 4k Instruct June 2024 better for coding?

Phi 3 Mini 4k Instruct June 2024 scores higher on coding benchmarks: 31.8 versus 29.3 in the Noometry coding category.

How many benchmarks do Command R and Phi 3 Mini 4k Instruct June 2024 share?

15 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Phi 3 Mini 4k Instruct June 2024 has 15.

Related comparisons

Go deeper