Model comparison

Qwen1.5 4b Chat vs Sonar

Sonar is the stronger model overall, scoring 38.5 to 28.8 on the Noometry Index.

Last verified . 0 shared benchmarks.

Qwen1.5 4b Chat Alibaba (Qwen)

28.8

Rank #322 Confirmed

Sonar Perplexity

38.5

Rank #187 Confirmed

Summary

  • The widest gap is in writing & preference, where Sonar leads 52.6 to 23.8.
  • Qwen1.5 4b Chat has downloadable open weights; the other is API-only.

Side by side

Qwen1.5 4b Chat and Sonar specifications
Qwen1.5 4b ChatSonar
ProviderAlibaba (Qwen)Perplexity
Noometry Index28.838.5
Released—2024-01-01
WeightsOpenProprietary
Context window—128K
Max output—4K
Input $ / M tokens—$1
Output $ / M tokens—$1
Results tracked137

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Sonar leads

Qwen1.5 4b Chat: 29.1 (#308), Sonar: 35.7 (#221)

Coding benchmarks
BenchmarkQwen1.5 4b ChatSonar
LiveBench Coding—35.1%
LMArena Coding999—

Reasoning Sonar leads

Qwen1.5 4b Chat: 18.5 (#279), Sonar: 21.1 (#227)

Reasoning benchmarks
BenchmarkQwen1.5 4b ChatSonar
LiveBench Reasoning—46.3%
LMArena Hard Prompts976—
LiveBench Data Analysis—37.9%
LiveBench—46.9%

Math Sonar leads

Qwen1.5 4b Chat: 30.4 (#234), Sonar: 33.7 (#200)

Math benchmarks
BenchmarkQwen1.5 4b ChatSonar
LiveBench Math—41.6%
LMArena Math1026—

Knowledge Not comparable

Qwen1.5 4b Chat: 26.7 (#255), Sonar: —

Knowledge benchmarks
BenchmarkQwen1.5 4b ChatSonar
LMArena Expert980—

Multilingual Not comparable

Qwen1.5 4b Chat: 24.1 (#290), Sonar: —

Multilingual benchmarks
BenchmarkQwen1.5 4b ChatSonar
LMArena Non-English979—
LMArena Chinese1024—
LMArena German902—
LMArena Russian952—

Instruction Following Sonar leads

Qwen1.5 4b Chat: 49.0 (#300), Sonar: 71.4 (#150)

Instruction Following benchmarks
BenchmarkQwen1.5 4b ChatSonar
LiveBench Instruction Following—76.2%
LMArena Instruction Following978—

Long Context Not comparable

Qwen1.5 4b Chat: 30.1 (#290), Sonar: —

Long Context benchmarks
BenchmarkQwen1.5 4b ChatSonar
LMArena Longer Query988—

Writing & Preference Sonar leads

Qwen1.5 4b Chat: 23.8 (#309), Sonar: 52.6 (#167)

Writing & Preference benchmarks
BenchmarkQwen1.5 4b ChatSonar
LMArena Text997—
LMArena Creative Writing969—
LMArena Multi-Turn977—
LiveBench Language—44.1%

Frequently asked questions

Is Qwen1.5 4b Chat better than Sonar?

Sonar is the stronger model overall, scoring 38.5 to 28.8 on the Noometry Index.

Is Qwen1.5 4b Chat or Sonar better for coding?

Sonar scores higher on coding benchmarks: 35.7 versus 29.1 in the Noometry coding category.

How many benchmarks do Qwen1.5 4b Chat and Sonar share?

0 benchmarks have published results for both models. Qwen1.5 4b Chat has 13 scored results on Noometry and Sonar has 7.

Related comparisons

Go deeper