Model comparison

Qwen2.5-Max vs Solar Pro4

Solar Pro4 is the stronger model overall, scoring 42.1 to 40.7 on the Noometry Index.

Last verified . 17 shared benchmarks.

Qwen2.5-Max Alibaba (Qwen)

40.7

Rank #146 Confirmed

Solar Pro4 Upstage

42.1

Rank #121 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Qwen2.5-Max scores higher in 1 category and Solar Pro4 in 7 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Solar Pro4 leads 39.8 to 35.3.

Side by side

Qwen2.5-Max and Solar Pro4 specifications
Qwen2.5-MaxSolar Pro4
ProviderAlibaba (Qwen)Upstage
Noometry Index40.742.1
Released2025-01-252026-08-06
WeightsProprietaryProprietary
Context window—524K
Max output—131K
Input $ / M tokens—$0.30
Output $ / M tokens—$1.20
Results tracked2718

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen2.5-Max leads

Qwen2.5-Max: 41.8 (#117), Solar Pro4: 40.1 (#149)

Coding benchmarks
BenchmarkQwen2.5-MaxSolar Pro4
LMArena Coding13591437
LMArena WebDev—1371
LiveBench Coding64.4%—

Reasoning Solar Pro4 leads

Qwen2.5-Max: 25.6 (#147), Solar Pro4: 28.5 (#104)

Reasoning benchmarks
BenchmarkQwen2.5-MaxSolar Pro4
LMArena Hard Prompts13601399
LiveBench Reasoning51.4%—
LiveBench Data Analysis67.9%—
Epoch Capabilities Index132.53—
LiveBench62.3%—

Math Solar Pro4 leads

Qwen2.5-Max: 36.9 (#162), Solar Pro4: 38.8 (#128)

Math benchmarks
BenchmarkQwen2.5-MaxSolar Pro4
LMArena Math13691416
LiveBench Math58.4%—

Knowledge Solar Pro4 leads

Qwen2.5-Max: 35.3 (#186), Solar Pro4: 39.8 (#129)

Knowledge benchmarks
BenchmarkQwen2.5-MaxSolar Pro4
LMArena Expert13371427
Confabulations21.8%—

Multilingual Too close to call

Qwen2.5-Max: 48.1 (#146), Solar Pro4: 48.7 (#139)

Multilingual benchmarks
BenchmarkQwen2.5-MaxSolar Pro4
LMArena Non-English13521361
LMArena Chinese13821415
LMArena French13961397
LMArena German13501364
LMArena Japanese13001309
LMArena Korean13041382
LMArena Russian13531360
LMArena Spanish13771401

Instruction Following Solar Pro4 leads

Qwen2.5-Max: 71.3 (#152), Solar Pro4: 72.7 (#132)

Instruction Following benchmarks
BenchmarkQwen2.5-MaxSolar Pro4
LMArena Instruction Following13351377
LiveBench Instruction Following75.3%—

Long Context Too close to call

Qwen2.5-Max: 41.4 (#142), Solar Pro4: 42.1 (#130)

Long Context benchmarks
BenchmarkQwen2.5-MaxSolar Pro4
LMArena Longer Query13581381

Writing & Preference Solar Pro4 leads

Qwen2.5-Max: 55.4 (#146), Solar Pro4: 56.5 (#138)

Writing & Preference benchmarks
BenchmarkQwen2.5-MaxSolar Pro4
LMArena Text13671386
LMArena Creative Writing13391316
LMArena Multi-Turn13641385
Short-Story Creative Writing72.9%—
LiveBench Language56.3%—

Frequently asked questions

Is Qwen2.5-Max better than Solar Pro4?

Solar Pro4 is the stronger model overall, scoring 42.1 to 40.7 on the Noometry Index.

Is Qwen2.5-Max or Solar Pro4 better for coding?

Qwen2.5-Max scores higher on coding benchmarks: 41.8 versus 40.1 in the Noometry coding category.

How many benchmarks do Qwen2.5-Max and Solar Pro4 share?

17 benchmarks have published results for both models. Qwen2.5-Max has 27 scored results on Noometry and Solar Pro4 has 18.

Related comparisons

Go deeper