Alibaba (Qwen), open weights

# QwQ-32B

> QwQ-32B by Alibaba (Qwen), released November 2024. Ranked #159 of 354 with a Noometry Index of 39.8. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/qwq-32b
- Last updated: 2026-10-10
- Title: QwQ-32B Benchmarks, Price & Rank (October 2026) | Noometry

QwQ-32B by Alibaba (Qwen) ranks 159th of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.8. Its strongest category is long context, where it ranks 11th.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #159 of 354
- **Index score:** 39.8
- **Evidence:** Confirmed 36 results
- **Provider:** [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba)
- **Released:** November 28, 2024
- **Weights:** Open weights
- **Reasoning:** Unknown
- **Context window:** —
- **Max output:** —
- **Input price:** Not listed
- **Output price:** Not listed
- **Blended price:** Not listed
- **Output speed:** Not measured
- **Value:** Not ranked
- **Knowledge cutoff:** Unknown

## Category scores

Each category score combines every public result we have in that category.

QwQ-32B category scores

1.  Coding 35.4
2.  Reasoning 23.7
3.  Math 38.0
4.  Knowledge 37.2
5.  Multilingual 44.8
6.  Instruction Following 72.6
7.  Long Context 49.0
8.  Writing & Preference 50.6
9.  020406080

QwQ-32B category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 35.4 | #226 | 5 |
| [Reasoning](https://noometry.com/best/reasoning) | 23.7 | #174 | 4 |
| [Math](https://noometry.com/best/math) | 38.0 | #143 | 3 |
| [Knowledge](https://noometry.com/best/knowledge) | 37.2 | #158 | 3 |
| [Multilingual](https://noometry.com/best/multilingual) | 44.8 | #176 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 72.6 | #137 | 2 |
| [Long Context](https://noometry.com/best/long-context) | 49.0 | #11 | 2 |
| [Writing & Preference](https://noometry.com/best/writing) | 50.6 | #180 | 6 |

## Strengths and weaknesses

Categories where QwQ-32B places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

QwQ-32B: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Long Context](https://noometry.com/best/long-context) | 49.0 | +8.0 | #11 of 296, top 4% |
| [Math](https://noometry.com/best/math) | 38.0 | +1.4 | #143 of 327, top 44% |
| [Instruction Following](https://noometry.com/best/instruction-following) | 72.6 | +1.3 | #137 of 305, top 45% |

### Weakest categories

QwQ-32B: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 35.4 | −3.4 | #226 of 340, top 67% |
| [Multilingual](https://noometry.com/best/multilingual) | 44.8 | −2.6 | #176 of 297, top 60% |
| [Writing & Preference](https://noometry.com/best/writing) | 50.6 | −3.1 | #180 of 312, top 58% |

## Closest competitors

The models ranked just above and below QwQ-32B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to QwQ-32B
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [Nemotron 3.5 Lightning](https://noometry.com/models/nemotron-3-5-lightning) | #155 | 40.0 | $0.0875 | — | [Compare](https://noometry.com/compare/nemotron-3-5-lightning-vs-qwq-32b) |
| [Qwen3.7 Flash](https://noometry.com/models/qwen3-7-flash) | #156 | 39.9 | $0.055 | — | [Compare](https://noometry.com/compare/qwen3-7-flash-vs-qwq-32b) |
| [Grok 3](https://noometry.com/models/grok-3) | #157 | 39.9 | — | 42 | [Compare](https://noometry.com/compare/grok-3-vs-qwq-32b) |
| [GLM-4.5V](https://noometry.com/models/glm-4-5v) | #158 | 39.8 | $0.90 | 34 | [Compare](https://noometry.com/compare/glm-4-5v-vs-qwq-32b) |
| [Step 1o Turbo 202506](https://noometry.com/models/step-1o-turbo-202506) | #160 | 39.7 | — | — | [Compare](https://noometry.com/compare/qwq-32b-vs-step-1o-turbo-202506) |
| [Nova 2 Lite](https://noometry.com/models/nova-2-lite) | #161 | 39.7 | $0.85 | — | [Compare](https://noometry.com/compare/nova-2-lite-vs-qwq-32b) |
| [DeepSeek-V3.2-Speciale](https://noometry.com/models/deepseek-v3-2-speciale) | #162 | 39.7 | $0.85 | — | [Compare](https://noometry.com/compare/deepseek-v3-2-speciale-vs-qwq-32b) |
| [Hunyuan Turbo 0110](https://noometry.com/models/hunyuan-turbo) | #163 | 39.6 | — | — | [Compare](https://noometry.com/compare/hunyuan-turbo-vs-qwq-32b) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

QwQ-32B Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Aider Polyglot](https://noometry.com/benchmarks/aider-polyglot) | 20.9% | #35 of 44, top 80% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [BigCodeBench Instruct](https://noometry.com/benchmarks/bigcodebench-instruct) | 44.6% | #22 of 64, top 35% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-11-28 |
| [LiveBench Coding](https://noometry.com/benchmarks/livebench-coding) | 37.2% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench Coding](https://noometry.com/benchmarks/livebench-coding) | 72.2% | #6 of 39, top 16% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1155 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1333 | #178 of 294, top 61% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [BigCodeBench Complete](https://noometry.com/benchmarks/bigcodebench-complete) | 54.4% | #22 of 66, top 34% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-11-28 |

### Reasoning

QwQ-32B Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Chess Puzzles](https://noometry.com/benchmarks/chess-puzzles) | 5% | #96 of 129, top 75% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-28 |
| [LiveBench Reasoning](https://noometry.com/benchmarks/livebench-reasoning) | 83.5% | #6 of 39, top 16% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench Reasoning](https://noometry.com/benchmarks/livebench-reasoning) | 57.7% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1325 | #177 of 297, top 60% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1163 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LiveBench Data Analysis](https://noometry.com/benchmarks/livebench-data-analysis) | 31.6% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench Data Analysis](https://noometry.com/benchmarks/livebench-data-analysis) | 65% | #11 of 39, top 29% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Epoch Capabilities Index](https://noometry.com/benchmarks/epoch-capabilities-index) | 137.6 | #117 of 213, top 55% |  | [Epoch AI](https://epoch.ai/eci) | 2025-03-05 |
| [ForecastBench](https://noometry.com/benchmarks/forecastbench) | 58.3 | #50 of 72, top 70% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench](https://noometry.com/benchmarks/livebench) | 40.3% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench](https://noometry.com/benchmarks/livebench) | 72% | #6 of 39, top 16% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Math

QwQ-32B Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [OTIS Mock AIME 2024-2025](https://noometry.com/benchmarks/otis-mock-aime) | 59.2% | #103 of 173, top 60% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-28 |
| [LiveBench Math](https://noometry.com/benchmarks/livebench-math) | 58.3% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench Math](https://noometry.com/benchmarks/livebench-math) | 77.8% | #6 of 39, top 16% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1213 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1359 | #157 of 285, top 56% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Knowledge

QwQ-32B Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [GPQA Diamond](https://noometry.com/benchmarks/gpqa-diamond) | 65.3% | #108 of 186, top 59% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-28 |
| [Confabulations](https://noometry.com/benchmarks/confabulations) (lower is better) | 15.6% | #19 of 51, top 38% |  | [Lech Mazur benchmarks](https://github.com/lechmazur/confabulations) |  |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1135 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1324 | #167 of 273, top 62% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Multilingual

QwQ-32B Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1130 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1305 | #176 of 297, top 60% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1218 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1378 | #153 of 285, top 54% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena French](https://noometry.com/benchmarks/arena-french) | 1336 | #148 of 223, top 67% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1313 | #146 of 231, top 64% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Japanese](https://noometry.com/benchmarks/arena-japanese) | 1262 | #137 of 211, top 65% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Korean](https://noometry.com/benchmarks/arena-korean) | 1279 | #139 of 213, top 66% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1297 | #178 of 283, top 63% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1115 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Spanish](https://noometry.com/benchmarks/arena-spanish) | 1354 | #142 of 226, top 63% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

QwQ-32B Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LiveBench Instruction Following](https://noometry.com/benchmarks/livebench-if) | 81.8% | #6 of 39, top 16% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench Instruction Following](https://noometry.com/benchmarks/livebench-if) | 35.6% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1156 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1297 | #183 of 298, top 62% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

QwQ-32B Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Fiction.LiveBench](https://noometry.com/benchmarks/fiction-livebench) | 83.3% | #11 of 47, top 24% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1308 | #182 of 291, top 63% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1167 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

QwQ-32B Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1329 | #174 of 297, top 59% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1162 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1288 | #178 of 295, top 61% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1134 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [Short-Story Creative Writing](https://noometry.com/benchmarks/lech-mazur-writing) | 80.2% | #15 of 39, top 39% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [EQ-Bench Creative Writing](https://noometry.com/benchmarks/eqbench-creative-writing) | 1257 | #83 of 115, top 73% |  | [EQ-Bench](https://eqbench.com/creative_writing.html) |  |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1141 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1314 | #180 of 295, top 62% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LiveBench Language](https://noometry.com/benchmarks/livebench-language) | 51.4% | #8 of 39, top 21% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench Language](https://noometry.com/benchmarks/livebench-language) | 21.1% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

## Compare QwQ-32B

-   [QwQ-32B vs GLM-4.5V](https://noometry.com/compare/glm-4-5v-vs-qwq-32b)
-   [QwQ-32B vs Step 1o Turbo 202506](https://noometry.com/compare/qwq-32b-vs-step-1o-turbo-202506)
-   [QwQ-32B vs Grok 3](https://noometry.com/compare/grok-3-vs-qwq-32b)
-   [QwQ-32B vs Nova 2 Lite](https://noometry.com/compare/nova-2-lite-vs-qwq-32b)
-   [QwQ-32B vs Qwen3.7 Flash](https://noometry.com/compare/qwen3-7-flash-vs-qwq-32b)
-   [QwQ-32B vs DeepSeek-V3.2-Speciale](https://noometry.com/compare/deepseek-v3-2-speciale-vs-qwq-32b)
-   [QwQ-32B vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-qwq-32b)
-   [QwQ-32B vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-qwq-32b)
-   [QwQ-32B vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-qwq-32b)
-   [QwQ-32B vs Kimi K3](https://noometry.com/compare/kimi-k3-vs-qwq-32b)
-   [QwQ-32B vs Grok 4.6](https://noometry.com/compare/grok-4-6-vs-qwq-32b)
-   [QwQ-32B vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-qwq-32b)
-   [QwQ-32B vs Muse Spark 1.3](https://noometry.com/compare/muse-spark-1-3-vs-qwq-32b)
-   [QwQ-32B vs DeepSeek V4 Pro](https://noometry.com/compare/deepseek-v4-pro-vs-qwq-32b)

## Other Alibaba (Qwen) models

-   [Qwen3.8 Max](https://noometry.com/models/qwen3-8-max)56.8
-   [Qwen3.7 Max](https://noometry.com/models/qwen3-7-max)51.5
-   [Qwen3.6 Max Preview](https://noometry.com/models/qwen3-6-max-preview)51.5
-   [Qwen3.6 Plus](https://noometry.com/models/qwen3-6-plus)47.5
-   [Qwen3.5 397B-A17B](https://noometry.com/models/qwen3-5-397b-a17b)46.0
-   [Qwen3.8 27B](https://noometry.com/models/qwen3-8-27b)46.0
-   [Qwen3.5 Max Preview](https://noometry.com/models/qwen3-5-max-preview)45.3
-   [Qwen3.7 Plus](https://noometry.com/models/qwen3-7-plus)45.3

## Frequently asked questions

### How good is QwQ-32B?

QwQ-32B by Alibaba (Qwen) ranks 159th of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.8. Its strongest category is long context, where it ranks 11th.

### Is QwQ-32B open source?

Yes. QwQ-32B's weights are downloadable; check the license for commercial terms.

### What are QwQ-32B's strengths and weaknesses?

Relative to other ranked models, QwQ-32B places best in long context, math, instruction following and lowest in coding, multilingual, writing & preference.

### What is QwQ-32B best at?

Its best category is long context, where it ranks 11th on Noometry.

### Cite this page

Noometry. (2026). QwQ-32B benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/qwq-32b

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/qwq-32b.md).
