Model comparison

o3-mini vs Solar Pro4

Solar Pro4 is the stronger model overall, scoring 42.1 to 36.7 on the Noometry Index.

Last verified . 17 shared benchmarks.

o3-mini OpenAI

36.7

Rank #212 Confirmed

Solar Pro4 Upstage

42.1

Rank #121 Confirmed

Summary

  • They share 17 benchmarks with published results for both. o3-mini scores higher in 2 categories and Solar Pro4 in 6 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Solar Pro4 leads 28.5 to 16.3.
  • Solar Pro4 is cheaper at $0.30 / $1.20 per million input/output tokens, against $1.10 / $4.40 for o3-mini.
  • Solar Pro4 accepts more context: 524K tokens versus 200K.

Side by side

o3-mini and Solar Pro4 specifications
o3-miniSolar Pro4
ProviderOpenAIUpstage
Noometry Index36.742.1
Released2024-12-202026-08-06
WeightsProprietaryProprietary
Context window200K524K
Max output100K131K
Input $ / M tokens$1.10$0.30
Output $ / M tokens$4.40$1.20
Results tracked5118

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

o3-mini: 40.8 (#132), Solar Pro4: 40.1 (#149)

Coding benchmarks
Benchmarko3-miniSolar Pro4
LMArena Coding13781437
Aider Polyglot60.4%—
LMArena WebDev—1371
SciCode39.8%—
GSO1.3%—
WeirdML43.7%—
LiveBench Coding82.7%—
CadEval54%—

Agentic & Tool Use Not comparable

o3-mini: 29.6 (#84), Solar Pro4: —

Agentic & Tool Use benchmarks
Benchmarko3-miniSolar Pro4
Cybench22.5%—

Reasoning Solar Pro4 leads

o3-mini: 16.3 (#305), Solar Pro4: 28.5 (#104)

Reasoning benchmarks
Benchmarko3-miniSolar Pro4
LMArena Hard Prompts13661399
ARC-AGI-23%—
SimpleBench22.8%—
ARC-AGI-134.5%—
CritPt0.3%—
Chess Puzzles17%—
LiveBench Reasoning89.6%—
Mystery Game Puzzles7%—
DTBench68.8%—
LiveBench Data Analysis70.6%—
LMCA19%—
Epoch Capabilities Index140.34—
ForecastBench59.6—
LiveBench75.9%—

Math Solar Pro4 leads

o3-mini: 28.1 (#244), Solar Pro4: 38.8 (#128)

Knowledge Solar Pro4 leads

o3-mini: 38.3 (#146), Solar Pro4: 39.8 (#129)

Knowledge benchmarks
Benchmarko3-miniSolar Pro4
LMArena Expert13641427
GPQA Diamond77%—
SimpleQA Verified15.3%—
Confabulations17.9%—

Multilingual Solar Pro4 leads

o3-mini: 45.7 (#164), Solar Pro4: 48.7 (#139)

Multilingual benchmarks
Benchmarko3-miniSolar Pro4
LMArena Non-English13191361
LMArena Chinese13791415
LMArena French13341397
LMArena German13031364
LMArena Japanese12861309
LMArena Korean13141382
LMArena Russian13041360
LMArena Spanish13211401

Instruction Following o3-mini leads

o3-mini: 75.1 (#72), Solar Pro4: 72.7 (#132)

Instruction Following benchmarks
Benchmarko3-miniSolar Pro4
LMArena Instruction Following13371377
LiveBench Instruction Following84.4%—

Long Context Solar Pro4 leads

o3-mini: 33.8 (#256), Solar Pro4: 42.1 (#130)

Long Context benchmarks
Benchmarko3-miniSolar Pro4
LMArena Longer Query13431381
Fiction.LiveBench50%—

Writing & Preference Solar Pro4 leads

o3-mini: 50.3 (#182), Solar Pro4: 56.5 (#138)

Writing & Preference benchmarks
Benchmarko3-miniSolar Pro4
LMArena Text13371386
LMArena Creative Writing12861316
LMArena Multi-Turn13201385
Short-Story Creative Writing61.7%—
LiveBench Language50.7%—

Frequently asked questions

Is o3-mini better than Solar Pro4?

Solar Pro4 is the stronger model overall, scoring 42.1 to 36.7 on the Noometry Index.

Which is cheaper, o3-mini or Solar Pro4?

Solar Pro4 is cheaper. It lists at $0.30 per million input tokens and $1.20 per million output tokens; o3-mini lists at $1.10 and $4.40.

Is o3-mini or Solar Pro4 better for coding?

They score almost the same on coding (40.8 vs 40.1); test both on your own repository before choosing.

Which has the bigger context window?

Solar Pro4 does, with 524K tokens against 200K.

How many benchmarks do o3-mini and Solar Pro4 share?

17 benchmarks have published results for both models. o3-mini has 51 scored results on Noometry and Solar Pro4 has 18.

Related comparisons

Go deeper