Model comparison

Qwen3 32B vs Solar Pro4

Solar Pro4 is the stronger model overall, scoring 42.1 to 39.2 on the Noometry Index.

Last verified . 13 shared benchmarks.

Qwen3 32B Alibaba (Qwen)

39.2

Rank #172 Confirmed

Solar Pro4 Upstage

42.1

Rank #121 Confirmed

Summary

  • They share 13 benchmarks with published results for both. Qwen3 32B scores higher in 3 categories and Solar Pro4 in 5 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Solar Pro4 leads 28.5 to 20.2.
  • Solar Pro4 is cheaper at $0.30 / $1.20 per million input/output tokens, against $0.70 / $2.80 for Qwen3 32B.
  • Solar Pro4 accepts more context: 524K tokens versus 131K.
  • Qwen3 32B has downloadable open weights; the other is API-only.

Side by side

Qwen3 32B and Solar Pro4 specifications
Qwen3 32BSolar Pro4
ProviderAlibaba (Qwen)Upstage
Noometry Index39.242.1
Released2025-042026-08-06
WeightsOpenProprietary
Context window131K524K
Max output16K131K
Input $ / M tokens$0.70$0.30
Output $ / M tokens$2.80$1.20
Results tracked2618

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Solar Pro4 leads

Qwen3 32B: 37.7 (#190), Solar Pro4: 40.1 (#149)

Coding benchmarks
BenchmarkQwen3 32BSolar Pro4
LMArena Coding13581437
Aider Polyglot40%—
LMArena WebDev—1371
SciCode35.4%—

Agentic & Tool Use Not comparable

Qwen3 32B: 32.6 (#62), Solar Pro4: —

Agentic & Tool Use benchmarks
BenchmarkQwen3 32BSolar Pro4
Berkeley Function Calling Leaderboard48.7%—

Reasoning Solar Pro4 leads

Qwen3 32B: 20.2 (#241), Solar Pro4: 28.5 (#104)

Reasoning benchmarks
BenchmarkQwen3 32BSolar Pro4
LMArena Hard Prompts13341399
Kagi LLM Benchmark54.9%—
CritPt0.3%—
Chess Puzzles5%—
DTBench67.5%—
LMCA17.3%—
Epoch Capabilities Index138.51—

Math Too close to call

Qwen3 32B: 39.7 (#99), Solar Pro4: 38.8 (#128)

Math benchmarks
BenchmarkQwen3 32BSolar Pro4
LMArena Math13991416
OTIS Mock AIME 2024-202566.9%—

Knowledge Too close to call

Qwen3 32B: 40.0 (#125), Solar Pro4: 39.8 (#129)

Knowledge benchmarks
BenchmarkQwen3 32BSolar Pro4
LMArena Expert13621427
GPQA Diamond65.7%—
Vectara Hallucination Rate5.9%—

Multilingual Solar Pro4 leads

Qwen3 32B: 45.6 (#167), Solar Pro4: 48.7 (#139)

Multilingual benchmarks
BenchmarkQwen3 32BSolar Pro4
LMArena Non-English13171361
LMArena Chinese13571415
LMArena German13411364
LMArena Russian13111360
LMArena French—1397
LMArena Japanese—1309
LMArena Korean—1382
LMArena Spanish—1401

Instruction Following Solar Pro4 leads

Qwen3 32B: 68.9 (#179), Solar Pro4: 72.7 (#132)

Instruction Following benchmarks
BenchmarkQwen3 32BSolar Pro4
LMArena Instruction Following13051377

Long Context Qwen3 32B leads

Qwen3 32B: 43.8 (#87), Solar Pro4: 42.1 (#130)

Long Context benchmarks
BenchmarkQwen3 32BSolar Pro4
LMArena Longer Query13271381
Fiction.LiveBench74.2%—

Writing & Preference Solar Pro4 leads

Qwen3 32B: 52.9 (#163), Solar Pro4: 56.5 (#138)

Writing & Preference benchmarks
BenchmarkQwen3 32BSolar Pro4
LMArena Text13401386
LMArena Creative Writing12971316
LMArena Multi-Turn13311385

Frequently asked questions

Is Qwen3 32B better than Solar Pro4?

Solar Pro4 is the stronger model overall, scoring 42.1 to 39.2 on the Noometry Index.

Which is cheaper, Qwen3 32B or Solar Pro4?

Solar Pro4 is cheaper. It lists at $0.30 per million input tokens and $1.20 per million output tokens; Qwen3 32B lists at $0.70 and $2.80.

Is Qwen3 32B or Solar Pro4 better for coding?

Solar Pro4 scores higher on coding benchmarks: 40.1 versus 37.7 in the Noometry coding category.

Which has the bigger context window?

Solar Pro4 does, with 524K tokens against 131K.

How many benchmarks do Qwen3 32B and Solar Pro4 share?

13 benchmarks have published results for both models. Qwen3 32B has 26 scored results on Noometry and Solar Pro4 has 18.

Related comparisons

Go deeper