Model comparison

QwQ-32B vs Solar Pro4

Solar Pro4 is the stronger model overall, scoring 42.1 to 39.8 on the Noometry Index.

Last verified . 17 shared benchmarks.

QwQ-32B Alibaba (Qwen)

39.8

Rank #159 Confirmed

Solar Pro4 Upstage

42.1

Rank #121 Confirmed

Summary

  • They share 17 benchmarks with published results for both. QwQ-32B scores higher in 1 category and Solar Pro4 in 7 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in long context, where QwQ-32B leads 49.0 to 42.1.
  • QwQ-32B has downloadable open weights; the other is API-only.

Side by side

QwQ-32B and Solar Pro4 specifications
QwQ-32BSolar Pro4
ProviderAlibaba (Qwen)Upstage
Noometry Index39.842.1
Released2024-11-282026-08-06
WeightsOpenProprietary
Context window—524K
Max output—131K
Input $ / M tokens—$0.30
Output $ / M tokens—$1.20
Results tracked3618

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Solar Pro4 leads

QwQ-32B: 35.4 (#226), Solar Pro4: 40.1 (#149)

Coding benchmarks
BenchmarkQwQ-32BSolar Pro4
LMArena Coding13331437
Aider Polyglot20.9%—
LMArena WebDev—1371
BigCodeBench Instruct44.6%—
LiveBench Coding72.2%—
BigCodeBench Complete54.4%—

Reasoning Solar Pro4 leads

QwQ-32B: 23.7 (#174), Solar Pro4: 28.5 (#104)

Reasoning benchmarks
BenchmarkQwQ-32BSolar Pro4
LMArena Hard Prompts13251399
Chess Puzzles5%—
LiveBench Reasoning83.5%—
LiveBench Data Analysis65%—
Epoch Capabilities Index137.6—
ForecastBench58.3—
LiveBench72%—

Math Too close to call

QwQ-32B: 38.0 (#143), Solar Pro4: 38.8 (#128)

Math benchmarks
BenchmarkQwQ-32BSolar Pro4
LMArena Math13591416
OTIS Mock AIME 2024-202559.2%—
LiveBench Math77.8%—

Knowledge Solar Pro4 leads

QwQ-32B: 37.2 (#158), Solar Pro4: 39.8 (#129)

Knowledge benchmarks
BenchmarkQwQ-32BSolar Pro4
LMArena Expert13241427
GPQA Diamond65.3%—
Confabulations15.6%—

Multilingual Solar Pro4 leads

QwQ-32B: 44.8 (#176), Solar Pro4: 48.7 (#139)

Multilingual benchmarks
BenchmarkQwQ-32BSolar Pro4
LMArena Non-English13051361
LMArena Chinese13781415
LMArena French13361397
LMArena German13131364
LMArena Japanese12621309
LMArena Korean12791382
LMArena Russian12971360
LMArena Spanish13541401

Instruction Following Too close to call

QwQ-32B: 72.6 (#137), Solar Pro4: 72.7 (#132)

Instruction Following benchmarks
BenchmarkQwQ-32BSolar Pro4
LMArena Instruction Following12971377
LiveBench Instruction Following81.8%—

Long Context QwQ-32B leads

QwQ-32B: 49.0 (#11), Solar Pro4: 42.1 (#130)

Long Context benchmarks
BenchmarkQwQ-32BSolar Pro4
LMArena Longer Query13081381
Fiction.LiveBench83.3%—

Writing & Preference Solar Pro4 leads

QwQ-32B: 50.6 (#180), Solar Pro4: 56.5 (#138)

Writing & Preference benchmarks
BenchmarkQwQ-32BSolar Pro4
LMArena Text13291386
LMArena Creative Writing12881316
LMArena Multi-Turn13141385
Short-Story Creative Writing80.2%—
EQ-Bench Creative Writing1257—
LiveBench Language51.4%—

Frequently asked questions

Is QwQ-32B better than Solar Pro4?

Solar Pro4 is the stronger model overall, scoring 42.1 to 39.8 on the Noometry Index.

Is QwQ-32B or Solar Pro4 better for coding?

Solar Pro4 scores higher on coding benchmarks: 40.1 versus 35.4 in the Noometry coding category.

How many benchmarks do QwQ-32B and Solar Pro4 share?

17 benchmarks have published results for both models. QwQ-32B has 36 scored results on Noometry and Solar Pro4 has 18.

Related comparisons

Go deeper