Model comparison

Qwen3.5 Max Preview vs Wizardlm 13b

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 31.4 on the Noometry Index.

Last verified . 10 shared benchmarks.

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Wizardlm 13b Microsoft

31.4

Rank #274 Confirmed

Summary

  • They share 10 benchmarks with published results for both. Qwen3.5 Max Preview scores higher in 7 categories and Wizardlm 13b in 0 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen3.5 Max Preview leads 66.0 to 30.5.
  • Wizardlm 13b has downloadable open weights; the other is API-only.

Side by side

Qwen3.5 Max Preview and Wizardlm 13b specifications
Qwen3.5 Max PreviewWizardlm 13b
ProviderAlibaba (Qwen)Microsoft
Noometry Index45.331.4
Released——
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1710

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 44.0 (#77), Wizardlm 13b: 30.1 (#298)

Coding benchmarks
BenchmarkQwen3.5 Max PreviewWizardlm 13b
LMArena Coding14871035

Reasoning Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 30.8 (#84), Wizardlm 13b: 19.4 (#259)

Reasoning benchmarks
BenchmarkQwen3.5 Max PreviewWizardlm 13b
LMArena Hard Prompts14831018

Math Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 40.1 (#94), Wizardlm 13b: 30.2 (#238)

Math benchmarks
BenchmarkQwen3.5 Max PreviewWizardlm 13b
LMArena Math14741017

Knowledge Not comparable

Qwen3.5 Max Preview: 41.8 (#107), Wizardlm 13b: —

Knowledge benchmarks
BenchmarkQwen3.5 Max PreviewWizardlm 13b
LMArena Expert1489—

Multilingual Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 56.2 (#22), Wizardlm 13b: 27.1 (#277)

Multilingual benchmarks
BenchmarkQwen3.5 Max PreviewWizardlm 13b
LMArena Non-English14651034
LMArena Chinese15341023
LMArena French1484—
LMArena German1487—
LMArena Japanese1495—
LMArena Korean1438—
LMArena Russian1471—
LMArena Spanish1470—

Instruction Following Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 77.0 (#31), Wizardlm 13b: 53.5 (#285)

Instruction Following benchmarks
BenchmarkQwen3.5 Max PreviewWizardlm 13b
LMArena Instruction Following14671048

Long Context Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 45.2 (#45), Wizardlm 13b: 32.0 (#273)

Long Context benchmarks
BenchmarkQwen3.5 Max PreviewWizardlm 13b
LMArena Longer Query14761054

Writing & Preference Qwen3.5 Max Preview leads

Qwen3.5 Max Preview: 66.0 (#41), Wizardlm 13b: 30.5 (#287)

Writing & Preference benchmarks
BenchmarkQwen3.5 Max PreviewWizardlm 13b
LMArena Text14701077
LMArena Creative Writing14641091
LMArena Multi-Turn14781047

Frequently asked questions

Is Qwen3.5 Max Preview better than Wizardlm 13b?

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 31.4 on the Noometry Index.

Is Qwen3.5 Max Preview or Wizardlm 13b better for coding?

Qwen3.5 Max Preview scores higher on coding benchmarks: 44.0 versus 30.1 in the Noometry coding category.

How many benchmarks do Qwen3.5 Max Preview and Wizardlm 13b share?

10 benchmarks have published results for both models. Qwen3.5 Max Preview has 17 scored results on Noometry and Wizardlm 13b has 10.

Related comparisons

Go deeper