Model comparison

Command R vs Phi-3.5-MoE

Command R has enough public results to be ranked (#272); Phi-3.5-MoE does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

Command R Cohere

31.4

Rank #272 Confirmed

Phi-3.5-MoE Microsoft

—

Unranked

Side by side

Command R and Phi-3.5-MoE specifications
Command RPhi-3.5-MoE
ProviderCohereMicrosoft
Noometry Index31.4—
Released2024-08-302024-08-17
WeightsOpenOpen
Context window128K—
Max output4K—
Input $ / M tokens$0.15—
Output $ / M tokens$0.60—
Results tracked293

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Command R: 29.3 (#306), Phi-3.5-MoE: —

Coding benchmarks
BenchmarkCommand RPhi-3.5-MoE
BigCodeBench Instruct37.1%—
LiveBench Coding17.9%—
LMArena Coding1169—
BigCodeBench Complete45.2%—

Reasoning Not comparable

Command R: 13.8 (#331), Phi-3.5-MoE: —

Reasoning benchmarks
BenchmarkCommand RPhi-3.5-MoE
LiveBench Reasoning21.9%—
LMArena Hard Prompts1164—
DTBench46.4%—
LiveBench Data Analysis33.3%—
LMCA9.2%—
LiveBench27.5%—
PIQA—88.6%

Math Not comparable

Command R: 28.0 (#246), Phi-3.5-MoE: —

Math benchmarks
BenchmarkCommand RPhi-3.5-MoE
LiveBench Math19.4%—
LMArena Math1155—
GSM8K—88.7%

Knowledge Not comparable

Command R: 31.0 (#221), Phi-3.5-MoE: —

Knowledge benchmarks
BenchmarkCommand RPhi-3.5-MoE
LMArena Expert1138—
BoolQ—84.6%
MMLU65.2%—

Multilingual Not comparable

Command R: 35.7 (#245), Phi-3.5-MoE: —

Multilingual benchmarks
BenchmarkCommand RPhi-3.5-MoE
LMArena Non-English1174—
LMArena Chinese1182—
LMArena French1162—
LMArena German1176—
LMArena Japanese1143—
LMArena Korean1163—
LMArena Russian1174—
LMArena Spanish1151—

Instruction Following Not comparable

Command R: 58.1 (#261), Phi-3.5-MoE: —

Instruction Following benchmarks
BenchmarkCommand RPhi-3.5-MoE
LiveBench Instruction Following55.6%—
LMArena Instruction Following1167—

Long Context Not comparable

Command R: 36.3 (#231), Phi-3.5-MoE: —

Long Context benchmarks
BenchmarkCommand RPhi-3.5-MoE
LMArena Longer Query1198—

Writing & Preference Not comparable

Command R: 38.2 (#254), Phi-3.5-MoE: —

Writing & Preference benchmarks
BenchmarkCommand RPhi-3.5-MoE
LMArena Text1187—
LMArena Creative Writing1170—
LMArena Multi-Turn1163—
LiveBench Language16.7%—

Frequently asked questions

Is Command R better than Phi-3.5-MoE?

Command R has enough public results to be ranked (#272); Phi-3.5-MoE does not yet, so treat this comparison as directional.

How many benchmarks do Command R and Phi-3.5-MoE share?

0 benchmarks have published results for both models. Command R has 29 scored results on Noometry and Phi-3.5-MoE has 3.

Related comparisons

Go deeper