Model comparison

Llama 3-70B vs Solar Pro4

Solar Pro4 is the stronger model overall, scoring 42.1 to 28.8 on the Noometry Index.

Last verified . 17 shared benchmarks.

Llama 3-70B Meta

28.8

Rank #323 Confirmed

Solar Pro4 Upstage

42.1

Rank #121 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Llama 3-70B scores higher in 0 categories and Solar Pro4 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Solar Pro4 leads 38.8 to 12.8.
  • Llama 3-70B has downloadable open weights; the other is API-only.

Side by side

Llama 3-70B and Solar Pro4 specifications
Llama 3-70BSolar Pro4
ProviderMetaUpstage
Noometry Index28.842.1
Released2024-04-182026-08-06
WeightsOpenProprietary
Context window—524K
Max output—131K
Input $ / M tokens—$0.30
Output $ / M tokens—$1.20
Results tracked3118

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Solar Pro4 leads

Llama 3-70B: 35.8 (#218), Solar Pro4: 40.1 (#149)

Coding benchmarks
BenchmarkLlama 3-70BSolar Pro4
LMArena Coding12061437
LMArena WebDev—1371
BigCodeBench Instruct43.6%—
BigCodeBench Complete54.5%—
HumanEval+72%—
MBPP+69%—

Agentic & Tool Use Not comparable

Llama 3-70B: 21.1 (#139), Solar Pro4: —

Agentic & Tool Use benchmarks
BenchmarkLlama 3-70BSolar Pro4
Cybench5%—

Reasoning Solar Pro4 leads

Llama 3-70B: 18.0 (#288), Solar Pro4: 28.5 (#104)

Reasoning benchmarks
BenchmarkLlama 3-70BSolar Pro4
LMArena Hard Prompts11951399
Kagi LLM Benchmark35.1%—
DTBench54.2%—
Epoch Capabilities Index122.93—
ForecastBench57.1—
WinoGrande83.5%—

Math Solar Pro4 leads

Llama 3-70B: 12.8 (#305), Solar Pro4: 38.8 (#128)

Math benchmarks
BenchmarkLlama 3-70BSolar Pro4
LMArena Math12181416
OTIS Mock AIME 2024-20254.3%—
MATH Level 522.6%—

Knowledge Solar Pro4 leads

Llama 3-70B: 20.8 (#277), Solar Pro4: 39.8 (#129)

Knowledge benchmarks
BenchmarkLlama 3-70BSolar Pro4
LMArena Expert11491427
GPQA Diamond40.6%—
MMLU79.3%—

Multilingual Solar Pro4 leads

Llama 3-70B: 33.6 (#251), Solar Pro4: 48.7 (#139)

Multilingual benchmarks
BenchmarkLlama 3-70BSolar Pro4
LMArena Non-English11421361
LMArena Chinese11141415
LMArena French12321397
LMArena German11691364
LMArena Japanese10171309
LMArena Korean10171382
LMArena Russian11591360
LMArena Spanish12411401

Instruction Following Solar Pro4 leads

Llama 3-70B: 62.5 (#238), Solar Pro4: 72.7 (#132)

Instruction Following benchmarks
BenchmarkLlama 3-70BSolar Pro4
LMArena Instruction Following11941377

Long Context Solar Pro4 leads

Llama 3-70B: 35.6 (#240), Solar Pro4: 42.1 (#130)

Long Context benchmarks
BenchmarkLlama 3-70BSolar Pro4
LMArena Longer Query11741381

Writing & Preference Solar Pro4 leads

Llama 3-70B: 42.8 (#231), Solar Pro4: 56.5 (#138)

Writing & Preference benchmarks
BenchmarkLlama 3-70BSolar Pro4
LMArena Text12211386
LMArena Creative Writing12101316
LMArena Multi-Turn12231385

Frequently asked questions

Is Llama 3-70B better than Solar Pro4?

Solar Pro4 is the stronger model overall, scoring 42.1 to 28.8 on the Noometry Index.

Is Llama 3-70B or Solar Pro4 better for coding?

Solar Pro4 scores higher on coding benchmarks: 40.1 versus 35.8 in the Noometry coding category.

How many benchmarks do Llama 3-70B and Solar Pro4 share?

17 benchmarks have published results for both models. Llama 3-70B has 31 scored results on Noometry and Solar Pro4 has 18.

Related comparisons

Go deeper