Model comparison
Qwen1.5 4b Chat vs Sonar
Sonar is the stronger model overall, scoring 38.5 to 28.8 on the Noometry Index.
Last verified . 0 shared benchmarks.
Summary
- The widest gap is in writing & preference, where Sonar leads 52.6 to 23.8.
- Qwen1.5 4b Chat has downloadable open weights; the other is API-only.
Side by side
| Qwen1.5 4b Chat | Sonar | |
|---|---|---|
| Provider | Alibaba (Qwen) | Perplexity |
| Noometry Index | 28.8 | 38.5 |
| Released | — | 2024-01-01 |
| Weights | Open | Proprietary |
| Context window | — | 128K |
| Max output | — | 4K |
| Input $ / M tokens | — | $1 |
| Output $ / M tokens | — | $1 |
| Results tracked | 13 | 7 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Sonar leads
Qwen1.5 4b Chat: 29.1 (#308), Sonar: 35.7 (#221)
| Benchmark | Qwen1.5 4b Chat | Sonar |
|---|---|---|
| LiveBench Coding | — | 35.1% |
| LMArena Coding | 999 | — |
Reasoning Sonar leads
Qwen1.5 4b Chat: 18.5 (#279), Sonar: 21.1 (#227)
| Benchmark | Qwen1.5 4b Chat | Sonar |
|---|---|---|
| LiveBench Reasoning | — | 46.3% |
| LMArena Hard Prompts | 976 | — |
| LiveBench Data Analysis | — | 37.9% |
| LiveBench | — | 46.9% |
Math Sonar leads
Qwen1.5 4b Chat: 30.4 (#234), Sonar: 33.7 (#200)
| Benchmark | Qwen1.5 4b Chat | Sonar |
|---|---|---|
| LiveBench Math | — | 41.6% |
| LMArena Math | 1026 | — |
Knowledge Not comparable
Qwen1.5 4b Chat: 26.7 (#255), Sonar: —
| Benchmark | Qwen1.5 4b Chat | Sonar |
|---|---|---|
| LMArena Expert | 980 | — |
Multilingual Not comparable
Qwen1.5 4b Chat: 24.1 (#290), Sonar: —
| Benchmark | Qwen1.5 4b Chat | Sonar |
|---|---|---|
| LMArena Non-English | 979 | — |
| LMArena Chinese | 1024 | — |
| LMArena German | 902 | — |
| LMArena Russian | 952 | — |
Instruction Following Sonar leads
Qwen1.5 4b Chat: 49.0 (#300), Sonar: 71.4 (#150)
| Benchmark | Qwen1.5 4b Chat | Sonar |
|---|---|---|
| LiveBench Instruction Following | — | 76.2% |
| LMArena Instruction Following | 978 | — |
Long Context Not comparable
Qwen1.5 4b Chat: 30.1 (#290), Sonar: —
| Benchmark | Qwen1.5 4b Chat | Sonar |
|---|---|---|
| LMArena Longer Query | 988 | — |
Writing & Preference Sonar leads
Qwen1.5 4b Chat: 23.8 (#309), Sonar: 52.6 (#167)
| Benchmark | Qwen1.5 4b Chat | Sonar |
|---|---|---|
| LMArena Text | 997 | — |
| LMArena Creative Writing | 969 | — |
| LMArena Multi-Turn | 977 | — |
| LiveBench Language | — | 44.1% |
Frequently asked questions
Is Qwen1.5 4b Chat better than Sonar?
Sonar is the stronger model overall, scoring 38.5 to 28.8 on the Noometry Index.
Is Qwen1.5 4b Chat or Sonar better for coding?
Sonar scores higher on coding benchmarks: 35.7 versus 29.1 in the Noometry coding category.
How many benchmarks do Qwen1.5 4b Chat and Sonar share?
0 benchmarks have published results for both models. Qwen1.5 4b Chat has 13 scored results on Noometry and Sonar has 7.