Alibaba (Qwen), open weights

QwQ-32B

QwQ-32B by Alibaba (Qwen) ranks 159th of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.8. Its strongest category is long context, where it ranks 11th.

Last verified

Specifications

Noometry rank
#159 of 354
Index score
39.8
Evidence
Confirmed 36 results
Released
November 28, 2024
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

QwQ-32B category scores
  1. Coding 35.4
  2. Reasoning 23.7
  3. Math 38.0
  4. Knowledge 37.2
  5. Multilingual 44.8
  6. Instruction Following 72.6
  7. Long Context 49.0
  8. Writing & Preference 50.6
QwQ-32B category ranks
CategoryScoreRankResults
Coding35.4#2265
Reasoning23.7#1744
Math38.0#1433
Knowledge37.2#1583
Multilingual44.8#1761
Instruction Following72.6#1372
Long Context49.0#112
Writing & Preference50.6#1806

Strengths and weaknesses

Categories where QwQ-32B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

QwQ-32B: strongest categories
CategoryScorevs medianRank
Long Context49.0+8.0#11 of 296, top 4%
Math38.0+1.4#143 of 327, top 44%
Instruction Following72.6+1.3#137 of 305, top 45%

Weakest categories

QwQ-32B: weakest categories
CategoryScorevs medianRank
Coding35.4−3.4#226 of 340, top 67%
Multilingual44.8−2.6#176 of 297, top 60%
Writing & Preference50.6−3.1#180 of 312, top 58%

Closest competitors

The models ranked just above and below QwQ-32B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to QwQ-32B
ModelRankScoreBlended $/MSpeed
Nemotron 3.5 Lightning#15540.0$0.0875—Compare
Qwen3.7 Flash#15639.9$0.055—Compare
Grok 3#15739.9—42Compare
GLM-4.5V#15839.8$0.9034Compare
Step 1o Turbo 202506#16039.7——Compare
Nova 2 Lite#16139.7$0.85—Compare
DeepSeek-V3.2-Speciale#16239.7$0.85—Compare
Hunyuan Turbo 0110#16339.6——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

QwQ-32B Coding benchmark results
BenchmarkScorePositionSettingSourceDate
Aider Polyglot20.9%#35 of 44, top 80%Epoch AI
BigCodeBench Instruct44.6%#22 of 64, top 35%BigCodeBench2024-11-28
LiveBench Coding37.2%Epoch AI
LiveBench Coding72.2%#6 of 39, top 16%Epoch AI
LMArena Coding1155LMArena2026-10-08
LMArena Coding1333#178 of 294, top 61%LMArena2026-10-08
BigCodeBench Complete54.4%#22 of 66, top 34%BigCodeBench2024-11-28

Reasoning

QwQ-32B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Chess Puzzles5%#96 of 129, top 75%Epoch AI2026-08-28
LiveBench Reasoning83.5%#6 of 39, top 16%Epoch AI
LiveBench Reasoning57.7%Epoch AI
LMArena Hard Prompts1325#177 of 297, top 60%LMArena2026-10-08
LMArena Hard Prompts1163LMArena2026-10-08
LiveBench Data Analysis31.6%Epoch AI
LiveBench Data Analysis65%#11 of 39, top 29%Epoch AI
Epoch Capabilities Index137.6#117 of 213, top 55%Epoch AI2025-03-05
ForecastBench58.3#50 of 72, top 70%Epoch AI
LiveBench40.3%Epoch AI
LiveBench72%#6 of 39, top 16%Epoch AI

Math

QwQ-32B Math benchmark results
BenchmarkScorePositionSettingSourceDate
OTIS Mock AIME 2024-202559.2%#103 of 173, top 60%Epoch AI2026-08-28
LiveBench Math58.3%Epoch AI
LiveBench Math77.8%#6 of 39, top 16%Epoch AI
LMArena Math1213LMArena2026-10-08
LMArena Math1359#157 of 285, top 56%LMArena2026-10-08

Knowledge

QwQ-32B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond65.3%#108 of 186, top 59%Epoch AI2026-08-28
Confabulations (lower is better)15.6%#19 of 51, top 38%Lech Mazur benchmarks
LMArena Expert1135LMArena2026-10-08
LMArena Expert1324#167 of 273, top 62%LMArena2026-10-08

Multilingual

QwQ-32B Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1130LMArena2026-10-08
LMArena Non-English1305#176 of 297, top 60%LMArena2026-10-08
LMArena Chinese1218LMArena2026-10-08
LMArena Chinese1378#153 of 285, top 54%LMArena2026-10-08
LMArena French1336#148 of 223, top 67%LMArena2026-10-08
LMArena German1313#146 of 231, top 64%LMArena2026-10-08
LMArena Japanese1262#137 of 211, top 65%LMArena2026-10-08
LMArena Korean1279#139 of 213, top 66%LMArena2026-10-08
LMArena Russian1297#178 of 283, top 63%LMArena2026-10-08
LMArena Russian1115LMArena2026-10-08
LMArena Spanish1354#142 of 226, top 63%LMArena2026-10-08

Instruction Following

QwQ-32B Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LiveBench Instruction Following81.8%#6 of 39, top 16%Epoch AI
LiveBench Instruction Following35.6%Epoch AI
LMArena Instruction Following1156LMArena2026-10-08
LMArena Instruction Following1297#183 of 298, top 62%LMArena2026-10-08

Long Context

QwQ-32B Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
Fiction.LiveBench83.3%#11 of 47, top 24%Epoch AI
LMArena Longer Query1308#182 of 291, top 63%LMArena2026-10-08
LMArena Longer Query1167LMArena2026-10-08

Writing & Preference

QwQ-32B Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1329#174 of 297, top 59%LMArena2026-10-08
LMArena Text1162LMArena2026-10-08
LMArena Creative Writing1288#178 of 295, top 61%LMArena2026-10-08
LMArena Creative Writing1134LMArena2026-10-08
Short-Story Creative Writing80.2%#15 of 39, top 39%Epoch AI
EQ-Bench Creative Writing1257#83 of 115, top 73%EQ-Bench
LMArena Multi-Turn1141LMArena2026-10-08
LMArena Multi-Turn1314#180 of 295, top 62%LMArena2026-10-08
LiveBench Language51.4%#8 of 39, top 21%Epoch AI
LiveBench Language21.1%Epoch AI

Compare QwQ-32B

Other Alibaba (Qwen) models

Frequently asked questions

How good is QwQ-32B?

QwQ-32B by Alibaba (Qwen) ranks 159th of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.8. Its strongest category is long context, where it ranks 11th.

Is QwQ-32B open source?

Yes. QwQ-32B's weights are downloadable; check the license for commercial terms.

What are QwQ-32B's strengths and weaknesses?

Relative to other ranked models, QwQ-32B places best in long context, math, instruction following and lowest in coding, multilingual, writing & preference.

What is QwQ-32B best at?

Its best category is long context, where it ranks 11th on Noometry.