Model comparison

GPT-5.2 Pro vs Qwen3.6 Max Preview

GPT-5.2 Pro and Qwen3.6 Max Preview score almost the same on the Noometry Index (52.3 vs 51.5), so choose on price, context window or the category you care about most.

Last verified . 4 shared benchmarks.

GPT-5.2 Pro OpenAI

52.3

Rank #39 Reported

Qwen3.6 Max Preview Alibaba (Qwen)

51.5

Rank #43 Confirmed

Summary

  • They share 4 benchmarks with published results for both. GPT-5.2 Pro scores higher in 2 categories and Qwen3.6 Max Preview in 0 categories; 2 gaps are clear of the uncertainty.
  • The widest gap is in math, where GPT-5.2 Pro leads 65.3 to 54.1.
  • The biggest single-benchmark swing is SimpleBench: 57.4% for GPT-5.2 Pro and 63% for Qwen3.6 Max Preview.
  • Qwen3.6 Max Preview is cheaper at $1.30 / $7.80 per million input/output tokens, against $21 / $168 for GPT-5.2 Pro.
  • GPT-5.2 Pro accepts more context: 400K tokens versus 262K.

Side by side

GPT-5.2 Pro and Qwen3.6 Max Preview specifications
GPT-5.2 ProQwen3.6 Max Preview
ProviderOpenAIAlibaba (Qwen)
Noometry Index52.351.5
Released2025-12-112026-04-20
WeightsProprietaryProprietary
Context window400K262K
Max output128K66K
Input $ / M tokens$21$1.30
Output $ / M tokens$168$7.80
Results tracked829

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

GPT-5.2 Pro: —, Qwen3.6 Max Preview: 48.7 (#54)

Coding benchmarks
BenchmarkGPT-5.2 ProQwen3.6 Max Preview
SWE-bench Verified—76.7%
LMArena WebDev—1482
LMArena Coding—1471

Agentic & Tool Use Not comparable

GPT-5.2 Pro: —, Qwen3.6 Max Preview: —

Agentic & Tool Use benchmarks
BenchmarkGPT-5.2 ProQwen3.6 Max Preview
Vending-Bench 2—4,254

Reasoning GPT-5.2 Pro leads

GPT-5.2 Pro: 51.5 (#33), Qwen3.6 Max Preview: 41.7 (#53)

Reasoning benchmarks
BenchmarkGPT-5.2 ProQwen3.6 Max Preview
SimpleBench57.4%63%
NYT Connections (extended)79.3%74.1%
Epoch Capabilities Index155.4149.24
ARC-AGI-254.2%—
ARC-AGI-190.5%—
Chess Puzzles—20%
LMArena Hard Prompts—1457
Mystery Game Puzzles—19%
DTBench—87.2%
LMCA—42.5%

Math GPT-5.2 Pro leads

GPT-5.2 Pro: 65.3 (#29), Qwen3.6 Max Preview: 54.1 (#46)

Math benchmarks
BenchmarkGPT-5.2 ProQwen3.6 Max Preview
FrontierMath Tier 4 (v1)31.3%4.2%
FrontierMath (Tiers 1-3)74%—
FrontierMath Tier 446%—
OTIS Mock AIME 2024-2025—91.1%
LMArena Math—1465
FrontierMath (Feb 2025 set)—23.1%

Knowledge Not comparable

GPT-5.2 Pro: —, Qwen3.6 Max Preview: 57.6 (#39)

Knowledge benchmarks
BenchmarkGPT-5.2 ProQwen3.6 Max Preview
GPQA Diamond—87.4%
SimpleQA Verified—52%
LMArena Expert—1478

Multilingual Not comparable

GPT-5.2 Pro: —, Qwen3.6 Max Preview: 54.2 (#48)

Multilingual benchmarks
BenchmarkGPT-5.2 ProQwen3.6 Max Preview
LMArena Non-English—1437
LMArena Chinese—1487
LMArena French—1449
LMArena Russian—1445
LMArena Spanish—1454

Instruction Following Not comparable

GPT-5.2 Pro: —, Qwen3.6 Max Preview: 75.7 (#55)

Instruction Following benchmarks
BenchmarkGPT-5.2 ProQwen3.6 Max Preview
LMArena Instruction Following—1438

Long Context Not comparable

GPT-5.2 Pro: —, Qwen3.6 Max Preview: 44.6 (#61)

Long Context benchmarks
BenchmarkGPT-5.2 ProQwen3.6 Max Preview
LMArena Longer Query—1457

Writing & Preference Not comparable

GPT-5.2 Pro: —, Qwen3.6 Max Preview: 63.8 (#60)

Writing & Preference benchmarks
BenchmarkGPT-5.2 ProQwen3.6 Max Preview
LMArena Text—1447
LMArena Creative Writing—1435
LMArena Multi-Turn—1456

Frequently asked questions

Is GPT-5.2 Pro better than Qwen3.6 Max Preview?

GPT-5.2 Pro and Qwen3.6 Max Preview score almost the same on the Noometry Index (52.3 vs 51.5), so choose on price, context window or the category you care about most.

Which is cheaper, GPT-5.2 Pro or Qwen3.6 Max Preview?

Qwen3.6 Max Preview is cheaper. It lists at $1.30 per million input tokens and $7.80 per million output tokens; GPT-5.2 Pro lists at $21 and $168.

Which has the bigger context window?

GPT-5.2 Pro does, with 400K tokens against 262K.

How many benchmarks do GPT-5.2 Pro and Qwen3.6 Max Preview share?

4 benchmarks have published results for both models. GPT-5.2 Pro has 8 scored results on Noometry and Qwen3.6 Max Preview has 29.

Related comparisons

Go deeper