Model comparison

Qwen3.5 27B vs Qwen3.6 Plus

Qwen3.6 Plus is the stronger model overall, scoring 47.5 to 41.9 on the Noometry Index.

Last verified . 24 shared benchmarks.

Qwen3.5 27B Alibaba (Qwen)

41.9

Rank #127 Confirmed

Qwen3.6 Plus Alibaba (Qwen)

47.5

Rank #62 Confirmed

Summary

  • They share 24 benchmarks with published results for both. Qwen3.5 27B scores higher in 0 categories and Qwen3.6 Plus in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Qwen3.6 Plus leads 56.1 to 38.0.
  • The biggest single-benchmark swing is Thematic Generalization: 45.5% for Qwen3.5 27B and 59.5% for Qwen3.6 Plus.
  • Qwen3.5 27B is cheaper at $0.30 / $2.40 per million input/output tokens, against $0.50 / $3 for Qwen3.6 Plus.
  • Qwen3.6 Plus accepts more context: 1M tokens versus 262K.
  • Qwen3.5 27B has downloadable open weights; the other is API-only.

Side by side

Qwen3.5 27B and Qwen3.6 Plus specifications
Qwen3.5 27BQwen3.6 Plus
ProviderAlibaba (Qwen)Alibaba (Qwen)
Noometry Index41.947.5
Released2026-02-232026-03-31
WeightsOpenProprietary
Context window262K1M
Max output66K66K
Input $ / M tokens$0.30$0.50
Output $ / M tokens$2.40$3
Results tracked2837

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.6 Plus leads

Qwen3.5 27B: 38.9 (#168), Qwen3.6 Plus: 40.8 (#130)

Coding benchmarks
BenchmarkQwen3.5 27BQwen3.6 Plus
LMArena WebDev13581461
LMArena Coding14271467
ALE-Bench349.45670.15
SWE-bench Verified—57.9%
SciCode—40.7%
WeirdML39.5%—

Agentic & Tool Use Not comparable

Qwen3.5 27B: —, Qwen3.6 Plus: —

Agentic & Tool Use benchmarks
BenchmarkQwen3.5 27BQwen3.6 Plus
Vending-Bench 2201.985,115

Reasoning Qwen3.6 Plus leads

Qwen3.5 27B: 27.5 (#117), Qwen3.6 Plus: 29.3 (#93)

Reasoning benchmarks
BenchmarkQwen3.5 27BQwen3.6 Plus
NYT Connections (extended)47.9%60.3%
Thematic Generalization45.5%59.5%
LMArena Hard Prompts14141449
DTBench82.4%81.9%
LMCA34%33.1%
CritPt—2.9%
Chess Puzzles—17%
Mystery Game Puzzles—12%
Epoch Capabilities Index—147.65

Math Qwen3.6 Plus leads

Qwen3.5 27B: 38.8 (#127), Qwen3.6 Plus: 51.8 (#54)

Knowledge Qwen3.6 Plus leads

Qwen3.5 27B: 38.0 (#150), Qwen3.6 Plus: 56.1 (#45)

Knowledge benchmarks
BenchmarkQwen3.5 27BQwen3.6 Plus
LMArena Expert14281454
GPQA Diamond—88.4%
SimpleQA Verified—44.1%
Vectara Hallucination Rate12.1%—

Multimodal Not comparable

Qwen3.5 27B: 39.4 (#59), Qwen3.6 Plus: —

Multimodal benchmarks
BenchmarkQwen3.5 27BQwen3.6 Plus
LMArena Vision1241—

Multilingual Qwen3.6 Plus leads

Qwen3.5 27B: 50.8 (#115), Qwen3.6 Plus: 53.3 (#70)

Multilingual benchmarks
BenchmarkQwen3.5 27BQwen3.6 Plus
LMArena Non-English13901424
LMArena Chinese14781477
LMArena French14101455
LMArena German13931452
LMArena Japanese13451389
LMArena Korean13581379
LMArena Russian13901434
LMArena Spanish14071432

Instruction Following Qwen3.6 Plus leads

Qwen3.5 27B: 73.5 (#119), Qwen3.6 Plus: 75.0 (#74)

Instruction Following benchmarks
BenchmarkQwen3.5 27BQwen3.6 Plus
LMArena Instruction Following13931425

Long Context Qwen3.6 Plus leads

Qwen3.5 27B: 43.1 (#106), Qwen3.6 Plus: 45.2 (#49)

Long Context benchmarks
BenchmarkQwen3.5 27BQwen3.6 Plus
LMArena Longer Query14131439
CL-bench—20.3%

Writing & Preference Qwen3.6 Plus leads

Qwen3.5 27B: 59.3 (#111), Qwen3.6 Plus: 62.2 (#82)

Writing & Preference benchmarks
BenchmarkQwen3.5 27BQwen3.6 Plus
LMArena Text14091437
LMArena Creative Writing13621404
LMArena Multi-Turn14101438

Frequently asked questions

Is Qwen3.5 27B better than Qwen3.6 Plus?

Qwen3.6 Plus is the stronger model overall, scoring 47.5 to 41.9 on the Noometry Index.

Which is cheaper, Qwen3.5 27B or Qwen3.6 Plus?

Qwen3.5 27B is cheaper. It lists at $0.30 per million input tokens and $2.40 per million output tokens; Qwen3.6 Plus lists at $0.50 and $3.

Is Qwen3.5 27B or Qwen3.6 Plus better for coding?

Qwen3.6 Plus scores higher on coding benchmarks: 40.8 versus 38.9 in the Noometry coding category.

Which has the bigger context window?

Qwen3.6 Plus does, with 1M tokens against 262K.

How many benchmarks do Qwen3.5 27B and Qwen3.6 Plus share?

24 benchmarks have published results for both models. Qwen3.5 27B has 28 scored results on Noometry and Qwen3.6 Plus has 37.

Related comparisons

Go deeper