Alibaba (Qwen), open weights

Qwen3 32B

Qwen3 32B by Alibaba (Qwen) ranks 172nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.2. Its strongest category is agentic & tool use, where it ranks 62nd. API pricing starts at $0.70 per million input tokens and $2.80 per million output tokens, with a 131K-token context window.

Last verified

Specifications

Noometry rank
#172 of 354
Index score
39.2
Evidence
Confirmed 26 results
Released
April 1, 2025
Weights
Open weights
Reasoning
Yes
Context window
131K
Max output
16K
Input price
$0.70 / M
Output price
$2.80 / M
Blended price
$1.22 / M
Output speed
86 tokens/s Kagi
Value
#130 of 219
Knowledge cutoff
April 2025
Input
text
Hugging Face
Qwen/Qwen3-32B

Category scores

Each category score combines every public result we have in that category.

Qwen3 32B category scores
  1. Coding 37.7
  2. Agentic & Tool Use 32.6
  3. Reasoning 20.2
  4. Math 39.7
  5. Knowledge 40.0
  6. Multilingual 45.6
  7. Instruction Following 68.9
  8. Long Context 43.8
  9. Writing & Preference 52.9
Qwen3 32B category ranks
CategoryScoreRankResults
Coding37.7#1903
Agentic & Tool Use32.6#621
Reasoning20.2#2416
Math39.7#992
Knowledge40.0#1253
Multilingual45.6#1671
Instruction Following68.9#1791
Long Context43.8#872
Writing & Preference52.9#1633

Strengths and weaknesses

Categories where Qwen3 32B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Qwen3 32B: strongest categories
CategoryScorevs medianRank
Long Context43.8+2.8#87 of 296, top 30%
Math39.7+3.1#99 of 327, top 31%
Knowledge40.0+2.7#125 of 314, top 40%

Weakest categories

Qwen3 32B: weakest categories
CategoryScorevs medianRank
Reasoning20.2−3.4#241 of 350, top 69%
Instruction Following68.9−2.4#179 of 305, top 59%
Multilingual45.6−1.8#167 of 297, top 57%

Closest competitors

The models ranked just above and below Qwen3 32B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen3 32B
ModelRankScoreBlended $/MSpeed
Olmo 3.1 32b Instruct#16839.4——Compare
Granite 4.2 3b#16939.4——Compare
Gemini 2.5 Flash#17039.3$0.85152Compare
Step 2 16k Exp 202412#17139.2——Compare
Gemini 2.0 Pro#17339.1——Compare
Molmo 2 8b#17439.1——Compare
Mercury 2#17539.1$0.38—Compare
Mistral Large 3#17639.1$0.387Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Qwen3 32B Coding benchmark results
BenchmarkScorePositionSettingSourceDate
Aider Polyglot40%#28 of 44, top 64%Epoch AI
SciCode35.4%#99 of 121, top 82%Epoch AI
LMArena Coding1358#170 of 294, top 58%LMArena2026-10-08

Agentic & Tool Use

Qwen3 32B Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
Berkeley Function Calling Leaderboard48.7%#20 of 49, top 41%fcBerkeley Function Calling Leaderboard

Reasoning

Qwen3 32B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Kagi LLM Benchmark54.9%#53 of 99, top 54%Kagi LLM Benchmark
Kagi LLM Benchmark48.7%Kagi LLM Benchmark
CritPt0.3%#98 of 134, top 74%Epoch AI
Chess Puzzles5%#94 of 129, top 73%Epoch AI2026-08-28
Chess Puzzles1%noneEpoch AI2026-08-28
LMArena Hard Prompts1334#170 of 297, top 58%LMArena2026-10-08
DTBench67.5%#99 of 151, top 66%Epoch AI
LMCA17.3%#98 of 125, top 79%Epoch AI
Epoch Capabilities Index138.51#113 of 213, top 54%Epoch AI2025-04-29

Math

Qwen3 32B Math benchmark results
BenchmarkScorePositionSettingSourceDate
OTIS Mock AIME 2024-202566.9%#96 of 173, top 56%Epoch AI2026-08-30
OTIS Mock AIME 2024-202523.1%noneEpoch AI2026-08-30
LMArena Math1399#126 of 285, top 45%LMArena2026-10-08

Knowledge

Qwen3 32B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond65.7%#105 of 186, top 57%Epoch AI2026-08-28
GPQA Diamond54.1%noneEpoch AI2026-08-30
Vectara Hallucination Rate (lower is better)5.9%#21 of 96, top 22%Vectara Hallucination Leaderboard
LMArena Expert1362#146 of 273, top 54%LMArena2026-10-08

Multilingual

Qwen3 32B Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1317#167 of 297, top 57%LMArena2026-10-08
LMArena Chinese1357#163 of 285, top 58%LMArena2026-10-08
LMArena German1341#136 of 231, top 59%LMArena2026-10-08
LMArena Russian1311#171 of 283, top 61%LMArena2026-10-08

Instruction Following

Qwen3 32B Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1305#173 of 298, top 59%LMArena2026-10-08

Long Context

Qwen3 32B Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
Fiction.LiveBench74.2%#16 of 47, top 35%Epoch AI
LMArena Longer Query1327#165 of 291, top 57%LMArena2026-10-08

Writing & Preference

Qwen3 32B Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1340#163 of 297, top 55%LMArena2026-10-08
LMArena Creative Writing1297#167 of 295, top 57%LMArena2026-10-08
LMArena Multi-Turn1331#172 of 295, top 59%LMArena2026-10-08

API pricing by provider

Qwen3 32B API prices
RouteInput $/MOutput $/MCached input $/MChecked
alibaba$0.70$2.80—2026-10-10
bedrock$0.15$0.60—2026-10-10
deepinfra$0.08$0.28—2026-10-10
openrouter$0.08$0.28—2026-10-10

Compare Qwen3 32B

Other Alibaba (Qwen) models

Frequently asked questions

How good is Qwen3 32B?

Qwen3 32B by Alibaba (Qwen) ranks 172nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.2. Its strongest category is agentic & tool use, where it ranks 62nd. API pricing starts at $0.70 per million input tokens and $2.80 per million output tokens, with a 131K-token context window.

How much does Qwen3 32B cost?

Qwen3 32B costs $0.70 per million input tokens and $2.80 per million output tokens on Alibaba (Qwen)'s own API.

What is Qwen3 32B's context window?

Qwen3 32B accepts up to 131K tokens of input and can write up to 16K tokens in one response.

Is Qwen3 32B open source?

Yes. Qwen3 32B's weights are downloadable from Hugging Face (Qwen/Qwen3-32B); check the license for commercial terms.

How fast is Qwen3 32B?

Qwen3 32B generated about 86 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.

What are Qwen3 32B's strengths and weaknesses?

Relative to other ranked models, Qwen3 32B places best in long context, math, knowledge and lowest in reasoning, instruction following, multilingual.

What is Qwen3 32B best at?

Its best category is agentic & tool use, where it ranks 62nd on Noometry.