Model comparison

DeepSeek-V2.5 (Sep 2024) vs Nova Premier 1.0

DeepSeek-V2.5 (Sep 2024) and Nova Premier 1.0 score almost the same on the Noometry Index (37.6 vs 38.3), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

DeepSeek-V2.5 (Sep 2024) DeepSeek

37.6

Rank #200 Confirmed

Nova Premier 1.0 Amazon

38.3

Rank #189 Confirmed

Summary

  • DeepSeek-V2.5 (Sep 2024) has downloadable open weights; the other is API-only.

Side by side

DeepSeek-V2.5 (Sep 2024) and Nova Premier 1.0 specifications
DeepSeek-V2.5 (Sep 2024)Nova Premier 1.0
ProviderDeepSeekAmazon
Noometry Index37.638.3
Released2024-09-062025-04-30
WeightsOpenProprietary
Context window—1M
Max output—10K
Input $ / M tokens—$2.50
Output $ / M tokens—$12.50
Results tracked226

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

DeepSeek-V2.5 (Sep 2024): 31.7 (#281), Nova Premier 1.0: —

Coding benchmarks
BenchmarkDeepSeek-V2.5 (Sep 2024)Nova Premier 1.0
Aider Polyglot17.8%—
BigCodeBench Instruct48.6%—
LMArena Coding1309—
BigCodeBench Complete53.2%—
HumanEval+83.5%—
MBPP+74.1%—

Reasoning DeepSeek-V2.5 (Sep 2024) leads

DeepSeek-V2.5 (Sep 2024): 25.6 (#145), Nova Premier 1.0: 23.0 (#185)

Reasoning benchmarks
BenchmarkDeepSeek-V2.5 (Sep 2024)Nova Premier 1.0
Kagi LLM Benchmark—44.8%
LMArena Hard Prompts1289—

Math DeepSeek-V2.5 (Sep 2024) leads

DeepSeek-V2.5 (Sep 2024): 35.9 (#177), Nova Premier 1.0: 33.8 (#199)

Math benchmarks
BenchmarkDeepSeek-V2.5 (Sep 2024)Nova Premier 1.0
Omni-MATH—35%
LMArena Math1288—

Knowledge Nova Premier 1.0 leads

DeepSeek-V2.5 (Sep 2024): 34.8 (#193), Nova Premier 1.0: 35.8 (#180)

Knowledge benchmarks
BenchmarkDeepSeek-V2.5 (Sep 2024)Nova Premier 1.0
MMLU-Pro—72.6%
GPQA (HELM)—51.8%
LMArena Expert1266—

Multilingual Not comparable

DeepSeek-V2.5 (Sep 2024): 42.5 (#193), Nova Premier 1.0: —

Multilingual benchmarks
BenchmarkDeepSeek-V2.5 (Sep 2024)Nova Premier 1.0
LMArena Non-English1273—
LMArena Chinese1318—
LMArena French1289—
LMArena German1258—
LMArena Japanese1228—
LMArena Korean1209—
LMArena Russian1289—
LMArena Spanish1248—

Instruction Following Too close to call

DeepSeek-V2.5 (Sep 2024): 67.5 (#194), Nova Premier 1.0: 66.8 (#204)

Instruction Following benchmarks
BenchmarkDeepSeek-V2.5 (Sep 2024)Nova Premier 1.0
IFEval—80.3%
LMArena Instruction Following1280—

Long Context Not comparable

DeepSeek-V2.5 (Sep 2024): 39.5 (#174), Nova Premier 1.0: —

Long Context benchmarks
BenchmarkDeepSeek-V2.5 (Sep 2024)Nova Premier 1.0
LMArena Longer Query1301—

Writing & Preference Nova Premier 1.0 leads

DeepSeek-V2.5 (Sep 2024): 49.8 (#187), Nova Premier 1.0: 51.0 (#178)

Writing & Preference benchmarks
BenchmarkDeepSeek-V2.5 (Sep 2024)Nova Premier 1.0
LMArena Text1294—
LMArena Creative Writing1285—
WildBench—78.8%
LMArena Multi-Turn1297—

Frequently asked questions

Is DeepSeek-V2.5 (Sep 2024) better than Nova Premier 1.0?

DeepSeek-V2.5 (Sep 2024) and Nova Premier 1.0 score almost the same on the Noometry Index (37.6 vs 38.3), so choose on price, context window or the category you care about most.

How many benchmarks do DeepSeek-V2.5 (Sep 2024) and Nova Premier 1.0 share?

0 benchmarks have published results for both models. DeepSeek-V2.5 (Sep 2024) has 22 scored results on Noometry and Nova Premier 1.0 has 6.

Related comparisons

Go deeper