Model comparison

Qwen Max vs Sonar

Sonar is the stronger model overall, scoring 38.5 to 34.7 on the Noometry Index.

Last verified . 0 shared benchmarks.

Qwen Max Alibaba (Qwen)

34.7

Rank #230 Confirmed

Sonar Perplexity

38.5

Rank #187 Confirmed

Summary

  • The widest gap is in math, where Sonar leads 33.7 to 22.3.
  • Sonar is cheaper at $1 / $1 per million input/output tokens, against $1.60 / $6.40 for Qwen Max.
  • Sonar accepts more context: 128K tokens versus 33K.

Side by side

Qwen Max and Sonar specifications
Qwen MaxSonar
ProviderAlibaba (Qwen)Perplexity
Noometry Index34.738.5
Released2024-04-032024-01-01
WeightsProprietaryProprietary
Context window33K128K
Max output8K4K
Input $ / M tokens$1.60$1
Output $ / M tokens$6.40$1
Results tracked237

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Sonar leads

Qwen Max: 30.7 (#292), Sonar: 35.7 (#221)

Coding benchmarks
BenchmarkQwen MaxSonar
Aider Polyglot21.8%—
LiveBench Coding—35.1%
LMArena Coding1288—

Reasoning Qwen Max leads

Qwen Max: 25.1 (#151), Sonar: 21.1 (#227)

Reasoning benchmarks
BenchmarkQwen MaxSonar
LiveBench Reasoning—46.3%
LMArena Hard Prompts1269—
LiveBench Data Analysis—37.9%
LiveBench—46.9%

Math Sonar leads

Qwen Max: 22.3 (#276), Sonar: 33.7 (#200)

Math benchmarks
BenchmarkQwen MaxSonar
OTIS Mock AIME 2024-202516.1%—
LiveBench Math—41.6%
LMArena Math1275—
MATH Level 567.2%—
FrontierMath (Feb 2025 set)1%—

Knowledge Not comparable

Qwen Max: 30.3 (#228), Sonar: —

Knowledge benchmarks
BenchmarkQwen MaxSonar
GPQA Diamond56.1%—
LMArena Expert1248—

Multilingual Not comparable

Qwen Max: 41.8 (#202), Sonar: —

Multilingual benchmarks
BenchmarkQwen MaxSonar
LMArena Non-English1263—
LMArena Chinese1254—
LMArena French1330—
LMArena German1254—
LMArena Japanese1205—
LMArena Korean1142—
LMArena Russian1274—
LMArena Spanish1290—

Instruction Following Sonar leads

Qwen Max: 66.5 (#208), Sonar: 71.4 (#150)

Instruction Following benchmarks
BenchmarkQwen MaxSonar
LiveBench Instruction Following—76.2%
LMArena Instruction Following1262—

Long Context Not comparable

Qwen Max: 39.4 (#180), Sonar: —

Long Context benchmarks
BenchmarkQwen MaxSonar
Fiction.LiveBench66.7%—
LMArena Longer Query1288—

Writing & Preference Sonar leads

Qwen Max: 47.8 (#205), Sonar: 52.6 (#167)

Writing & Preference benchmarks
BenchmarkQwen MaxSonar
LMArena Text1282—
LMArena Creative Writing1248—
LMArena Multi-Turn1277—
LiveBench Language—44.1%

Frequently asked questions

Is Qwen Max better than Sonar?

Sonar is the stronger model overall, scoring 38.5 to 34.7 on the Noometry Index.

Which is cheaper, Qwen Max or Sonar?

Sonar is cheaper. It lists at $1 per million input tokens and $1 per million output tokens; Qwen Max lists at $1.60 and $6.40.

Is Qwen Max or Sonar better for coding?

Sonar scores higher on coding benchmarks: 35.7 versus 30.7 in the Noometry coding category.

Which has the bigger context window?

Sonar does, with 128K tokens against 33K.

How many benchmarks do Qwen Max and Sonar share?

0 benchmarks have published results for both models. Qwen Max has 23 scored results on Noometry and Sonar has 7.

Related comparisons

Go deeper