Model comparison

Qwen2.5 Plus 1127 vs Sonar

Qwen2.5 Plus 1127 and Sonar score almost the same on the Noometry Index (38.8 vs 38.5), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Qwen2.5 Plus 1127 Alibaba (Qwen)

38.8

Rank #181 Confirmed

Sonar Perplexity

38.5

Rank #187 Confirmed

Summary

  • The widest gap is in reasoning, where Qwen2.5 Plus 1127 leads 25.9 to 21.1.

Side by side

Qwen2.5 Plus 1127 and Sonar specifications
Qwen2.5 Plus 1127Sonar
ProviderAlibaba (Qwen)Perplexity
Noometry Index38.838.5
Released—2024-01-01
WeightsProprietaryProprietary
Context window—128K
Max output—4K
Input $ / M tokens—$1
Output $ / M tokens—$1
Results tracked147

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen2.5 Plus 1127 leads

Qwen2.5 Plus 1127: 38.5 (#175), Sonar: 35.7 (#221)

Coding benchmarks
BenchmarkQwen2.5 Plus 1127Sonar
LiveBench Coding—35.1%
LMArena Coding1314—

Reasoning Qwen2.5 Plus 1127 leads

Qwen2.5 Plus 1127: 25.9 (#141), Sonar: 21.1 (#227)

Reasoning benchmarks
BenchmarkQwen2.5 Plus 1127Sonar
LiveBench Reasoning—46.3%
LMArena Hard Prompts1299—
LiveBench Data Analysis—37.9%
LiveBench—46.9%

Math Qwen2.5 Plus 1127 leads

Qwen2.5 Plus 1127: 36.1 (#174), Sonar: 33.7 (#200)

Math benchmarks
BenchmarkQwen2.5 Plus 1127Sonar
LiveBench Math—41.6%
LMArena Math1298—

Knowledge Not comparable

Qwen2.5 Plus 1127: 35.5 (#183), Sonar: —

Knowledge benchmarks
BenchmarkQwen2.5 Plus 1127Sonar
LMArena Expert1289—

Multilingual Not comparable

Qwen2.5 Plus 1127: 41.9 (#201), Sonar: —

Multilingual benchmarks
BenchmarkQwen2.5 Plus 1127Sonar
LMArena Non-English1265—
LMArena Chinese1314—
LMArena German1231—
LMArena Japanese1207—
LMArena Russian1271—

Instruction Following Sonar leads

Qwen2.5 Plus 1127: 67.2 (#199), Sonar: 71.4 (#150)

Instruction Following benchmarks
BenchmarkQwen2.5 Plus 1127Sonar
LiveBench Instruction Following—76.2%
LMArena Instruction Following1275—

Long Context Not comparable

Qwen2.5 Plus 1127: 39.2 (#184), Sonar: —

Long Context benchmarks
BenchmarkQwen2.5 Plus 1127Sonar
LMArena Longer Query1292—

Writing & Preference Sonar leads

Qwen2.5 Plus 1127: 49.4 (#192), Sonar: 52.6 (#167)

Writing & Preference benchmarks
BenchmarkQwen2.5 Plus 1127Sonar
LMArena Text1299—
LMArena Creative Writing1262—
LMArena Multi-Turn1299—
LiveBench Language—44.1%

Frequently asked questions

Is Qwen2.5 Plus 1127 better than Sonar?

Qwen2.5 Plus 1127 and Sonar score almost the same on the Noometry Index (38.8 vs 38.5), so choose on price, context window or the category you care about most.

Is Qwen2.5 Plus 1127 or Sonar better for coding?

Qwen2.5 Plus 1127 scores higher on coding benchmarks: 38.5 versus 35.7 in the Noometry coding category.

How many benchmarks do Qwen2.5 Plus 1127 and Sonar share?

0 benchmarks have published results for both models. Qwen2.5 Plus 1127 has 14 scored results on Noometry and Sonar has 7.

Related comparisons

Go deeper