Model comparison

DeepSeek Coder 1.3B vs Qwen3.6 Max Preview

Qwen3.6 Max Preview has enough public results to be ranked (#43); DeepSeek Coder 1.3B does not yet, so treat this comparison as directional.

Last verified . 1 shared benchmarks.

DeepSeek Coder 1.3B DeepSeek

35.0

Unranked Sparse

Qwen3.6 Max Preview Alibaba (Qwen)

51.5

Rank #43 Confirmed

Summary

  • They share 1 benchmark with published results for both. DeepSeek Coder 1.3B scores higher in 0 categories and Qwen3.6 Max Preview in 1 category; one gap is clear of the uncertainty.
  • The widest gap is in coding, where Qwen3.6 Max Preview leads 48.7 to 31.2.
  • DeepSeek Coder 1.3B has downloadable open weights; the other is API-only.

Side by side

DeepSeek Coder 1.3B and Qwen3.6 Max Preview specifications
DeepSeek Coder 1.3BQwen3.6 Max Preview
ProviderDeepSeekAlibaba (Qwen)
Noometry Index35.051.5
Released2023-11-022026-04-20
WeightsOpenProprietary
Context window—262K
Max output—66K
Input $ / M tokens—$1.30
Output $ / M tokens—$7.80
Results tracked929

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.6 Max Preview leads

DeepSeek Coder 1.3B: 31.2 (#287), Qwen3.6 Max Preview: 48.7 (#54)

Coding benchmarks
BenchmarkDeepSeek Coder 1.3BQwen3.6 Max Preview
SWE-bench Verified—76.7%
LMArena WebDev—1482
BigCodeBench Instruct22.8%—
LMArena Coding—1471
BigCodeBench Complete29.6%—
HumanEval+60.4%—
MBPP+54.8%—

Agentic & Tool Use Not comparable

DeepSeek Coder 1.3B: —, Qwen3.6 Max Preview: —

Agentic & Tool Use benchmarks
BenchmarkDeepSeek Coder 1.3BQwen3.6 Max Preview
Vending-Bench 2—4,254

Reasoning Not comparable

DeepSeek Coder 1.3B: —, Qwen3.6 Max Preview: 41.7 (#53)

Reasoning benchmarks
BenchmarkDeepSeek Coder 1.3BQwen3.6 Max Preview
Epoch Capabilities Index63.6149.24
SimpleBench—63%
NYT Connections (extended)—74.1%
Chess Puzzles—20%
LMArena Hard Prompts—1457
Mystery Game Puzzles—19%
DTBench—87.2%
LMCA—42.5%
WinoGrande53.3%—

Math Not comparable

DeepSeek Coder 1.3B: —, Qwen3.6 Max Preview: 54.1 (#46)

Math benchmarks
BenchmarkDeepSeek Coder 1.3BQwen3.6 Max Preview
OTIS Mock AIME 2024-2025—91.1%
LMArena Math—1465
FrontierMath (Feb 2025 set)—23.1%
FrontierMath Tier 4 (v1)—4.2%
GSM8K4.4%—

Knowledge Not comparable

DeepSeek Coder 1.3B: —, Qwen3.6 Max Preview: 57.6 (#39)

Knowledge benchmarks
BenchmarkDeepSeek Coder 1.3BQwen3.6 Max Preview
GPQA Diamond—87.4%
SimpleQA Verified—52%
LMArena Expert—1478
ARC (AI2) Challenge25.4%—
MMLU25.8%—

Multilingual Not comparable

DeepSeek Coder 1.3B: —, Qwen3.6 Max Preview: 54.2 (#48)

Multilingual benchmarks
BenchmarkDeepSeek Coder 1.3BQwen3.6 Max Preview
LMArena Non-English—1437
LMArena Chinese—1487
LMArena French—1449
LMArena Russian—1445
LMArena Spanish—1454

Instruction Following Not comparable

DeepSeek Coder 1.3B: —, Qwen3.6 Max Preview: 75.7 (#55)

Instruction Following benchmarks
BenchmarkDeepSeek Coder 1.3BQwen3.6 Max Preview
LMArena Instruction Following—1438

Long Context Not comparable

DeepSeek Coder 1.3B: —, Qwen3.6 Max Preview: 44.6 (#61)

Long Context benchmarks
BenchmarkDeepSeek Coder 1.3BQwen3.6 Max Preview
LMArena Longer Query—1457

Writing & Preference Not comparable

DeepSeek Coder 1.3B: —, Qwen3.6 Max Preview: 63.8 (#60)

Writing & Preference benchmarks
BenchmarkDeepSeek Coder 1.3BQwen3.6 Max Preview
LMArena Text—1447
LMArena Creative Writing—1435
LMArena Multi-Turn—1456

Frequently asked questions

Is DeepSeek Coder 1.3B better than Qwen3.6 Max Preview?

Qwen3.6 Max Preview has enough public results to be ranked (#43); DeepSeek Coder 1.3B does not yet, so treat this comparison as directional.

Is DeepSeek Coder 1.3B or Qwen3.6 Max Preview better for coding?

Qwen3.6 Max Preview scores higher on coding benchmarks: 48.7 versus 31.2 in the Noometry coding category.

How many benchmarks do DeepSeek Coder 1.3B and Qwen3.6 Max Preview share?

1 benchmark has published results for both models. DeepSeek Coder 1.3B has 9 scored results on Noometry and Qwen3.6 Max Preview has 29.

Related comparisons

Go deeper