Model comparison

DeepSeek LLM 67B vs Solar Pro4

Solar Pro4 is the stronger model overall, scoring 42.1 to 24.9 on the Noometry Index.

Last verified . 10 shared benchmarks.

DeepSeek LLM 67B DeepSeek

24.9

Rank #347 Confirmed

Solar Pro4 Upstage

42.1

Rank #121 Confirmed

Summary

  • They share 10 benchmarks with published results for both. DeepSeek LLM 67B scores higher in 0 categories and Solar Pro4 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Solar Pro4 leads 39.8 to 7.0.
  • DeepSeek LLM 67B has downloadable open weights; the other is API-only.

Side by side

DeepSeek LLM 67B and Solar Pro4 specifications
DeepSeek LLM 67BSolar Pro4
ProviderDeepSeekUpstage
Noometry Index24.942.1
Released2023-11-292026-08-06
WeightsOpenProprietary
Context window—524K
Max output—131K
Input $ / M tokens—$0.30
Output $ / M tokens—$1.20
Results tracked1518

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Solar Pro4 leads

DeepSeek LLM 67B: 31.9 (#278), Solar Pro4: 40.1 (#149)

Coding benchmarks
BenchmarkDeepSeek LLM 67BSolar Pro4
LMArena Coding10961437
LMArena WebDev—1371

Reasoning Solar Pro4 leads

DeepSeek LLM 67B: 16.5 (#304), Solar Pro4: 28.5 (#104)

Reasoning benchmarks
BenchmarkDeepSeek LLM 67BSolar Pro4
LMArena Hard Prompts10701399
Chess Puzzles0%—
Epoch Capabilities Index110.5—

Math Solar Pro4 leads

DeepSeek LLM 67B: 8.7 (#324), Solar Pro4: 38.8 (#128)

Math benchmarks
BenchmarkDeepSeek LLM 67BSolar Pro4
LMArena Math11081416
OTIS Mock AIME 2024-20250.8%—
MATH Level 56.4%—

Knowledge Solar Pro4 leads

DeepSeek LLM 67B: 7.0 (#313), Solar Pro4: 39.8 (#129)

Knowledge benchmarks
BenchmarkDeepSeek LLM 67BSolar Pro4
GPQA Diamond24.6%—
LMArena Expert—1427

Multilingual Solar Pro4 leads

DeepSeek LLM 67B: 29.4 (#267), Solar Pro4: 48.7 (#139)

Multilingual benchmarks
BenchmarkDeepSeek LLM 67BSolar Pro4
LMArena Non-English10731361
LMArena Chinese11321415
LMArena French—1397
LMArena German—1364
LMArena Japanese—1309
LMArena Korean—1382
LMArena Russian—1360
LMArena Spanish—1401

Instruction Following Solar Pro4 leads

DeepSeek LLM 67B: 55.4 (#277), Solar Pro4: 72.7 (#132)

Instruction Following benchmarks
BenchmarkDeepSeek LLM 67BSolar Pro4
LMArena Instruction Following10791377

Long Context Solar Pro4 leads

DeepSeek LLM 67B: 33.1 (#265), Solar Pro4: 42.1 (#130)

Long Context benchmarks
BenchmarkDeepSeek LLM 67BSolar Pro4
LMArena Longer Query10921381

Writing & Preference Solar Pro4 leads

DeepSeek LLM 67B: 31.6 (#282), Solar Pro4: 56.5 (#138)

Writing & Preference benchmarks
BenchmarkDeepSeek LLM 67BSolar Pro4
LMArena Text11051386
LMArena Creative Writing10671316
LMArena Multi-Turn10821385

Frequently asked questions

Is DeepSeek LLM 67B better than Solar Pro4?

Solar Pro4 is the stronger model overall, scoring 42.1 to 24.9 on the Noometry Index.

Is DeepSeek LLM 67B or Solar Pro4 better for coding?

Solar Pro4 scores higher on coding benchmarks: 40.1 versus 31.9 in the Noometry coding category.

How many benchmarks do DeepSeek LLM 67B and Solar Pro4 share?

10 benchmarks have published results for both models. DeepSeek LLM 67B has 15 scored results on Noometry and Solar Pro4 has 18.

Related comparisons

Go deeper