Model comparison

Qwen3.6 Plus vs Sonar

Qwen3.6 Plus is the stronger model overall, scoring 47.5 to 38.5 on the Noometry Index.

Last verified . 0 shared benchmarks.

Qwen3.6 Plus Alibaba (Qwen)

47.5

Rank #62 Confirmed

Sonar Perplexity

38.5

Rank #187 Confirmed

Summary

  • The widest gap is in math, where Qwen3.6 Plus leads 51.8 to 33.7.
  • Sonar is cheaper at $1 / $1 per million input/output tokens, against $0.50 / $3 for Qwen3.6 Plus.
  • Qwen3.6 Plus accepts more context: 1M tokens versus 128K.

Side by side

Qwen3.6 Plus and Sonar specifications
Qwen3.6 PlusSonar
ProviderAlibaba (Qwen)Perplexity
Noometry Index47.538.5
Released2026-03-312024-01-01
WeightsProprietaryProprietary
Context window1M128K
Max output66K4K
Input $ / M tokens$0.50$1
Output $ / M tokens$3$1
Results tracked377

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.6 Plus leads

Qwen3.6 Plus: 40.8 (#130), Sonar: 35.7 (#221)

Coding benchmarks
BenchmarkQwen3.6 PlusSonar
SWE-bench Verified57.9%—
LMArena WebDev1461—
SciCode40.7%—
LiveBench Coding—35.1%
LMArena Coding1467—
ALE-Bench670.15—

Agentic & Tool Use Not comparable

Qwen3.6 Plus: —, Sonar: —

Agentic & Tool Use benchmarks
BenchmarkQwen3.6 PlusSonar
Vending-Bench 25,115—

Reasoning Qwen3.6 Plus leads

Qwen3.6 Plus: 29.3 (#93), Sonar: 21.1 (#227)

Reasoning benchmarks
BenchmarkQwen3.6 PlusSonar
NYT Connections (extended)60.3%—
CritPt2.9%—
Chess Puzzles17%—
Thematic Generalization59.5%—
LiveBench Reasoning—46.3%
LMArena Hard Prompts1449—
Mystery Game Puzzles12%—
DTBench81.9%—
LiveBench Data Analysis—37.9%
LMCA33.1%—
Epoch Capabilities Index147.65—
LiveBench—46.9%

Math Qwen3.6 Plus leads

Qwen3.6 Plus: 51.8 (#54), Sonar: 33.7 (#200)

Math benchmarks
BenchmarkQwen3.6 PlusSonar
FrontierMath (Tiers 1-3)38.2%—
OTIS Mock AIME 2024-202593.3%—
LiveBench Math—41.6%
LMArena Math1450—
FrontierMath (Feb 2025 set)26.2%—
FrontierMath Tier 4 (v1)8.3%—

Knowledge Not comparable

Qwen3.6 Plus: 56.1 (#45), Sonar: —

Knowledge benchmarks
BenchmarkQwen3.6 PlusSonar
GPQA Diamond88.4%—
SimpleQA Verified44.1%—
LMArena Expert1454—

Multilingual Not comparable

Qwen3.6 Plus: 53.3 (#70), Sonar: —

Multilingual benchmarks
BenchmarkQwen3.6 PlusSonar
LMArena Non-English1424—
LMArena Chinese1477—
LMArena French1455—
LMArena German1452—
LMArena Japanese1389—
LMArena Korean1379—
LMArena Russian1434—
LMArena Spanish1432—

Instruction Following Qwen3.6 Plus leads

Qwen3.6 Plus: 75.0 (#74), Sonar: 71.4 (#150)

Instruction Following benchmarks
BenchmarkQwen3.6 PlusSonar
LiveBench Instruction Following—76.2%
LMArena Instruction Following1425—

Long Context Not comparable

Qwen3.6 Plus: 45.2 (#49), Sonar: —

Long Context benchmarks
BenchmarkQwen3.6 PlusSonar
CL-bench20.3%—
LMArena Longer Query1439—

Writing & Preference Qwen3.6 Plus leads

Qwen3.6 Plus: 62.2 (#82), Sonar: 52.6 (#167)

Writing & Preference benchmarks
BenchmarkQwen3.6 PlusSonar
LMArena Text1437—
LMArena Creative Writing1404—
LMArena Multi-Turn1438—
LiveBench Language—44.1%

Frequently asked questions

Is Qwen3.6 Plus better than Sonar?

Qwen3.6 Plus is the stronger model overall, scoring 47.5 to 38.5 on the Noometry Index.

Which is cheaper, Qwen3.6 Plus or Sonar?

Sonar is cheaper. It lists at $1 per million input tokens and $1 per million output tokens; Qwen3.6 Plus lists at $0.50 and $3.

Is Qwen3.6 Plus or Sonar better for coding?

Qwen3.6 Plus scores higher on coding benchmarks: 40.8 versus 35.7 in the Noometry coding category.

Which has the bigger context window?

Qwen3.6 Plus does, with 1M tokens against 128K.

How many benchmarks do Qwen3.6 Plus and Sonar share?

0 benchmarks have published results for both models. Qwen3.6 Plus has 37 scored results on Noometry and Sonar has 7.

Related comparisons

Go deeper