Model comparison

phi-3-medium 14B vs Sonar

Sonar is the stronger model overall, scoring 38.5 to 29.7 on the Noometry Index.

Last verified . 0 shared benchmarks.

phi-3-medium 14B Microsoft

29.7

Rank #306 Reported

Sonar Perplexity

38.5

Rank #187 Confirmed

Summary

  • The widest gap is in math, where Sonar leads 33.7 to 27.3.
  • phi-3-medium 14B has downloadable open weights; the other is API-only.

Side by side

phi-3-medium 14B and Sonar specifications
phi-3-medium 14BSonar
ProviderMicrosoftPerplexity
Noometry Index29.738.5
Released2024-04-232024-01-01
WeightsOpenProprietary
Context window—128K
Max output—4K
Input $ / M tokens—$1
Output $ / M tokens—$1
Results tracked137

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding phi-3-medium 14B leads

phi-3-medium 14B: 36.8 (#201), Sonar: 35.7 (#221)

Coding benchmarks
Benchmarkphi-3-medium 14BSonar
BigCodeBench Instruct37.6%—
LiveBench Coding—35.1%
BigCodeBench Complete48.7%—

Reasoning Not comparable

phi-3-medium 14B: —, Sonar: 21.1 (#227)

Reasoning benchmarks
Benchmarkphi-3-medium 14BSonar
LiveBench Reasoning—46.3%
LiveBench Data Analysis—37.9%
Adversarial NLI55.8%—
BIG-Bench Hard81.4%—
Epoch Capabilities Index121.23—
HellaSwag82.4%—
LiveBench—46.9%
WinoGrande81.5%—

Math Sonar leads

phi-3-medium 14B: 27.3 (#250), Sonar: 33.7 (#200)

Math benchmarks
Benchmarkphi-3-medium 14BSonar
LiveBench Math—41.6%
MATH Level 517.6%—

Knowledge Not comparable

phi-3-medium 14B: 9.1 (#306), Sonar: —

Knowledge benchmarks
Benchmarkphi-3-medium 14BSonar
GPQA Diamond27.6%—
ARC (AI2) Challenge91.6%—
MMLU78%—
OpenBookQA87.4%—
TriviaQA73.9%—

Instruction Following Not comparable

phi-3-medium 14B: —, Sonar: 71.4 (#150)

Instruction Following benchmarks
Benchmarkphi-3-medium 14BSonar
LiveBench Instruction Following—76.2%

Writing & Preference Not comparable

phi-3-medium 14B: —, Sonar: 52.6 (#167)

Writing & Preference benchmarks
Benchmarkphi-3-medium 14BSonar
LiveBench Language—44.1%

Frequently asked questions

Is phi-3-medium 14B better than Sonar?

Sonar is the stronger model overall, scoring 38.5 to 29.7 on the Noometry Index.

Is phi-3-medium 14B or Sonar better for coding?

phi-3-medium 14B scores higher on coding benchmarks: 36.8 versus 35.7 in the Noometry coding category.

How many benchmarks do phi-3-medium 14B and Sonar share?

0 benchmarks have published results for both models. phi-3-medium 14B has 13 scored results on Noometry and Sonar has 7.

Related comparisons

Go deeper