Model comparison

Qwen1.5-72B vs Sonar

Sonar is the stronger model overall, scoring 38.5 to 30.8 on the Noometry Index.

Last verified . 0 shared benchmarks.

Qwen1.5-72B Alibaba (Qwen)

30.8

Rank #285 Confirmed

Sonar Perplexity

38.5

Rank #187 Confirmed

Summary

  • The widest gap is in writing & preference, where Sonar leads 52.6 to 37.3.
  • Qwen1.5-72B has downloadable open weights; the other is API-only.

Side by side

Qwen1.5-72B and Sonar specifications
Qwen1.5-72BSonar
ProviderAlibaba (Qwen)Perplexity
Noometry Index30.838.5
Released2024-02-042024-01-01
WeightsOpenProprietary
Context window—128K
Max output—4K
Input $ / M tokens—$1
Output $ / M tokens—$1
Results tracked227

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Sonar leads

Qwen1.5-72B: 31.9 (#277), Sonar: 35.7 (#221)

Coding benchmarks
BenchmarkQwen1.5-72BSonar
BigCodeBench Instruct33.2%—
LiveBench Coding—35.1%
LMArena Coding1165—
BigCodeBench Complete40.3%—
HumanEval+59.1%—
MBPP+61.6%—

Reasoning Qwen1.5-72B leads

Qwen1.5-72B: 22.2 (#203), Sonar: 21.1 (#227)

Reasoning benchmarks
BenchmarkQwen1.5-72BSonar
LiveBench Reasoning—46.3%
LMArena Hard Prompts1148—
LiveBench Data Analysis—37.9%
LiveBench—46.9%

Math Too close to call

Qwen1.5-72B: 33.2 (#205), Sonar: 33.7 (#200)

Math benchmarks
BenchmarkQwen1.5-72BSonar
LiveBench Math—41.6%
LMArena Math1164—

Knowledge Not comparable

Qwen1.5-72B: 11.5 (#300), Sonar: —

Knowledge benchmarks
BenchmarkQwen1.5-72BSonar
GPQA Diamond28.8%—
LMArena Expert1136—

Multilingual Not comparable

Qwen1.5-72B: 33.2 (#253), Sonar: —

Multilingual benchmarks
BenchmarkQwen1.5-72BSonar
LMArena Non-English1135—
LMArena Chinese1186—
LMArena French1159—
LMArena German1084—
LMArena Japanese1061—
LMArena Korean1050—
LMArena Russian1104—
LMArena Spanish1110—

Instruction Following Sonar leads

Qwen1.5-72B: 59.3 (#256), Sonar: 71.4 (#150)

Instruction Following benchmarks
BenchmarkQwen1.5-72BSonar
LiveBench Instruction Following—76.2%
LMArena Instruction Following1141—

Long Context Not comparable

Qwen1.5-72B: 35.1 (#243), Sonar: —

Long Context benchmarks
BenchmarkQwen1.5-72BSonar
LMArena Longer Query1157—

Writing & Preference Sonar leads

Qwen1.5-72B: 37.3 (#258), Sonar: 52.6 (#167)

Writing & Preference benchmarks
BenchmarkQwen1.5-72BSonar
LMArena Text1166—
LMArena Creative Writing1137—
LMArena Multi-Turn1160—
LiveBench Language—44.1%

Frequently asked questions

Is Qwen1.5-72B better than Sonar?

Sonar is the stronger model overall, scoring 38.5 to 30.8 on the Noometry Index.

Is Qwen1.5-72B or Sonar better for coding?

Sonar scores higher on coding benchmarks: 35.7 versus 31.9 in the Noometry coding category.

How many benchmarks do Qwen1.5-72B and Sonar share?

0 benchmarks have published results for both models. Qwen1.5-72B has 22 scored results on Noometry and Sonar has 7.

Related comparisons

Go deeper