Model comparison

Mistral Large vs Solar Pro4

Solar Pro4 is the stronger model overall, scoring 42.1 to 31.9 on the Noometry Index.

Last verified . 17 shared benchmarks.

Mistral Large Mistral AI

31.9

Rank #263 Confirmed

Solar Pro4 Upstage

42.1

Rank #121 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Mistral Large scores higher in 0 categories and Solar Pro4 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Solar Pro4 leads 38.8 to 18.2.
  • Solar Pro4 is cheaper at $0.30 / $1.20 per million input/output tokens, against $2 / $6 for Mistral Large.
  • Solar Pro4 accepts more context: 524K tokens versus 131K.
  • Mistral Large has downloadable open weights; the other is API-only.

Side by side

Mistral Large and Solar Pro4 specifications
Mistral LargeSolar Pro4
ProviderMistral AIUpstage
Noometry Index31.942.1
Released2024-02-262026-08-06
WeightsOpenProprietary
Context window131K524K
Max output16K131K
Input $ / M tokens$2$0.30
Output $ / M tokens$6$1.20
Results tracked5118

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Solar Pro4 leads

Mistral Large: 34.3 (#240), Solar Pro4: 40.1 (#149)

Coding benchmarks
BenchmarkMistral LargeSolar Pro4
LMArena Coding12771437
LMArena WebDev—1371
SciCode36.2%—
BigCodeBench Instruct30%—
LiveBench Coding47.1%—
BigCodeBench Complete38.3%—
ALE-Bench264.7—
HumanEval+62.2%—
MBPP+59.5%—

Agentic & Tool Use Not comparable

Mistral Large: 28.6 (#89), Solar Pro4: —

Agentic & Tool Use benchmarks
BenchmarkMistral LargeSolar Pro4
Berkeley Function Calling Leaderboard38.4%—

Reasoning Solar Pro4 leads

Mistral Large: 15.8 (#310), Solar Pro4: 28.5 (#104)

Reasoning benchmarks
BenchmarkMistral LargeSolar Pro4
LMArena Hard Prompts12571399
SimpleBench22.5%—
CritPt0%—
LiveBench Reasoning43.5%—
DTBench65.1%—
LiveBench Data Analysis50.1%—
LMCA16.7%—
Epoch Capabilities Index128.52—
ForecastBench57.1—
LiveBench48.4%—

Math Solar Pro4 leads

Mistral Large: 18.2 (#291), Solar Pro4: 38.8 (#128)

Math benchmarks
BenchmarkMistral LargeSolar Pro4
LMArena Math12621416
OTIS Mock AIME 2024-20258.5%—
Omni-MATH28.1%—
LiveBench Math42.5%—
MATH Level 550.3%—
FrontierMath (Feb 2025 set)0.3%—

Knowledge Solar Pro4 leads

Mistral Large: 30.1 (#230), Solar Pro4: 39.8 (#129)

Knowledge benchmarks
BenchmarkMistral LargeSolar Pro4
LMArena Expert12321427
GPQA Diamond51.3%—
MMLU-Pro59.9%—
Confabulations21.4%—
Vectara Hallucination Rate4.5%—
GPQA (HELM)43.5%—
MMLU80%—

Multilingual Solar Pro4 leads

Mistral Large: 40.0 (#219), Solar Pro4: 48.7 (#139)

Multilingual benchmarks
BenchmarkMistral LargeSolar Pro4
LMArena Non-English12371361
LMArena Chinese12401415
LMArena French13251397
LMArena German12541364
LMArena Japanese11881309
LMArena Korean12021382
LMArena Russian12571360
LMArena Spanish12681401

Instruction Following Solar Pro4 leads

Mistral Large: 67.9 (#191), Solar Pro4: 72.7 (#132)

Instruction Following benchmarks
BenchmarkMistral LargeSolar Pro4
LMArena Instruction Following12491377
LiveBench Instruction Following67.9%—
IFEval87.7%—

Long Context Solar Pro4 leads

Mistral Large: 38.3 (#199), Solar Pro4: 42.1 (#130)

Long Context benchmarks
BenchmarkMistral LargeSolar Pro4
LMArena Longer Query12611381

Writing & Preference Solar Pro4 leads

Mistral Large: 40.7 (#242), Solar Pro4: 56.5 (#138)

Writing & Preference benchmarks
BenchmarkMistral LargeSolar Pro4
LMArena Text12661386
LMArena Creative Writing12431316
LMArena Multi-Turn12601385
Short-Story Creative Writing69%—
EQ-Bench Creative Writing985—
WildBench80.1%—
LiveBench Language39.4%—

Frequently asked questions

Is Mistral Large better than Solar Pro4?

Solar Pro4 is the stronger model overall, scoring 42.1 to 31.9 on the Noometry Index.

Which is cheaper, Mistral Large or Solar Pro4?

Solar Pro4 is cheaper. It lists at $0.30 per million input tokens and $1.20 per million output tokens; Mistral Large lists at $2 and $6.

Is Mistral Large or Solar Pro4 better for coding?

Solar Pro4 scores higher on coding benchmarks: 40.1 versus 34.3 in the Noometry coding category.

Which has the bigger context window?

Solar Pro4 does, with 524K tokens against 131K.

How many benchmarks do Mistral Large and Solar Pro4 share?

17 benchmarks have published results for both models. Mistral Large has 51 scored results on Noometry and Solar Pro4 has 18.

Related comparisons

Go deeper