Model comparison

Qwen2-72B vs Sonar

Sonar is the stronger model overall, scoring 38.5 to 30.0 on the Noometry Index.

Last verified . 0 shared benchmarks.

Qwen2-72B Alibaba (Qwen)

30.0

Rank #300 Confirmed

Sonar Perplexity

38.5

Rank #187 Confirmed

Summary

  • The widest gap is in writing & preference, where Sonar leads 52.6 to 40.8.
  • Qwen2-72B has downloadable open weights; the other is API-only.

Side by side

Qwen2-72B and Sonar specifications
Qwen2-72BSonar
ProviderAlibaba (Qwen)Perplexity
Noometry Index30.038.5
Released2024-06-072024-01-01
WeightsOpenProprietary
Context window—128K
Max output—4K
Input $ / M tokens—$1
Output $ / M tokens—$1
Results tracked267

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Sonar leads

Qwen2-72B: 29.1 (#310), Sonar: 35.7 (#221)

Coding benchmarks
BenchmarkQwen2-72BSonar
WeirdML11.3%—
BigCodeBench Instruct38.5%—
LiveBench Coding—35.1%
LMArena Coding1196—
BigCodeBench Complete54%—

Agentic & Tool Use Not comparable

Qwen2-72B: 17.0 (#146), Sonar: —

Agentic & Tool Use benchmarks
BenchmarkQwen2-72BSonar
TheAgentCompany1.1%—
METR Time Horizons29.9%—

Reasoning Qwen2-72B leads

Qwen2-72B: 23.2 (#181), Sonar: 21.1 (#227)

Reasoning benchmarks
BenchmarkQwen2-72BSonar
LiveBench Reasoning—46.3%
LMArena Hard Prompts1191—
LiveBench Data Analysis—37.9%
Epoch Capabilities Index125.28—
LiveBench—46.9%

Math Sonar leads

Qwen2-72B: 30.2 (#236), Sonar: 33.7 (#200)

Math benchmarks
BenchmarkQwen2-72BSonar
LiveBench Math—41.6%
LMArena Math1235—
MATH Level 539.1%—

Knowledge Not comparable

Qwen2-72B: 21.2 (#275), Sonar: —

Knowledge benchmarks
BenchmarkQwen2-72BSonar
GPQA Diamond40.8%—
LMArena Expert1171—
MMLU82.4%—

Multilingual Not comparable

Qwen2-72B: 35.9 (#244), Sonar: —

Multilingual benchmarks
BenchmarkQwen2-72BSonar
LMArena Non-English1176—
LMArena Chinese1240—
LMArena French1170—
LMArena German1151—
LMArena Japanese1111—
LMArena Korean1083—
LMArena Russian1169—
LMArena Spanish1169—

Instruction Following Sonar leads

Qwen2-72B: 61.7 (#241), Sonar: 71.4 (#150)

Instruction Following benchmarks
BenchmarkQwen2-72BSonar
LiveBench Instruction Following—76.2%
LMArena Instruction Following1181—

Long Context Not comparable

Qwen2-72B: 36.1 (#235), Sonar: —

Long Context benchmarks
BenchmarkQwen2-72BSonar
LMArena Longer Query1192—

Writing & Preference Sonar leads

Qwen2-72B: 40.8 (#241), Sonar: 52.6 (#167)

Writing & Preference benchmarks
BenchmarkQwen2-72BSonar
LMArena Text1203—
LMArena Creative Writing1181—
LMArena Multi-Turn1196—
LiveBench Language—44.1%

Frequently asked questions

Is Qwen2-72B better than Sonar?

Sonar is the stronger model overall, scoring 38.5 to 30.0 on the Noometry Index.

Is Qwen2-72B or Sonar better for coding?

Sonar scores higher on coding benchmarks: 35.7 versus 29.1 in the Noometry coding category.

How many benchmarks do Qwen2-72B and Sonar share?

0 benchmarks have published results for both models. Qwen2-72B has 26 scored results on Noometry and Sonar has 7.

Related comparisons

Go deeper