Model comparison

Seed 2.0 Pro vs Dolly 2.0-12b

Seed 2.0 Pro is the stronger model overall, scoring 43.2 to 25.5 on the Noometry Index.

Last verified . 9 shared benchmarks.

Seed 2.0 Pro ByteDance Seed

43.2

Rank #96 Confirmed

Dolly 2.0-12b Databricks

25.5

Rank #342 Confirmed

Summary

  • They share 9 benchmarks with published results for both. Seed 2.0 Pro scores higher in 6 categories and Dolly 2.0-12b in 0 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Seed 2.0 Pro leads 62.9 to 15.2.
  • Dolly 2.0-12b has downloadable open weights; the other is API-only.

Side by side

Seed 2.0 Pro and Dolly 2.0-12b specifications
Seed 2.0 ProDolly 2.0-12b
ProviderByteDance SeedDatabricks
Noometry Index43.225.5
Released2026-02-142023-04-11
WeightsProprietaryOpen
Context window256K—
Max output128K—
Input $ / M tokens$0.50—
Output $ / M tokens$3—
Results tracked2017

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Seed 2.0 Pro leads

Seed 2.0 Pro: 43.5 (#86), Dolly 2.0-12b: 23.4 (#332)

Coding benchmarks
BenchmarkSeed 2.0 ProDolly 2.0-12b
LMArena Coding1472776

Reasoning Seed 2.0 Pro leads

Seed 2.0 Pro: 24.1 (#165), Dolly 2.0-12b: 15.3 (#316)

Reasoning benchmarks
BenchmarkSeed 2.0 ProDolly 2.0-12b
LMArena Hard Prompts1453804
NYT Connections (extended)28.4%—
Thematic Generalization57.1%—
Epoch Capabilities Index—89.67
HellaSwag—70.8%
PIQA—75.4%
WinoGrande—61.8%

Math Seed 2.0 Pro leads

Seed 2.0 Pro: 39.3 (#108), Dolly 2.0-12b: 27.3 (#251)

Math benchmarks
BenchmarkSeed 2.0 ProDolly 2.0-12b
LMArena Math1439871

Knowledge Not comparable

Seed 2.0 Pro: 40.2 (#122), Dolly 2.0-12b: —

Knowledge benchmarks
BenchmarkSeed 2.0 ProDolly 2.0-12b
LMArena Expert1440—
ARC (AI2) Challenge—39.6%
BoolQ—56.3%
MMLU—26.2%
OpenBookQA—39.2%

Multimodal Not comparable

Seed 2.0 Pro: 41.5 (#35), Dolly 2.0-12b: —

Multimodal benchmarks
BenchmarkSeed 2.0 ProDolly 2.0-12b
LMArena Vision1274—

Multilingual Seed 2.0 Pro leads

Seed 2.0 Pro: 54.5 (#39), Dolly 2.0-12b: 17.4 (#296)

Multilingual benchmarks
BenchmarkSeed 2.0 ProDolly 2.0-12b
LMArena Non-English1441836
LMArena Chinese1489836
LMArena French1471—
LMArena German1442—
LMArena Japanese1408—
LMArena Korean1411—
LMArena Russian1449—
LMArena Spanish1460—

Instruction Following Seed 2.0 Pro leads

Seed 2.0 Pro: 74.5 (#91), Dolly 2.0-12b: 38.7 (#304)

Instruction Following benchmarks
BenchmarkSeed 2.0 ProDolly 2.0-12b
LMArena Instruction Following1414814

Long Context Not comparable

Seed 2.0 Pro: 43.6 (#90), Dolly 2.0-12b: —

Long Context benchmarks
BenchmarkSeed 2.0 ProDolly 2.0-12b
LMArena Longer Query1428—

Writing & Preference Seed 2.0 Pro leads

Seed 2.0 Pro: 62.9 (#69), Dolly 2.0-12b: 15.2 (#311)

Writing & Preference benchmarks
BenchmarkSeed 2.0 ProDolly 2.0-12b
LMArena Text1448851
LMArena Creative Writing1406864
LMArena Multi-Turn1441740

Frequently asked questions

Is Seed 2.0 Pro better than Dolly 2.0-12b?

Seed 2.0 Pro is the stronger model overall, scoring 43.2 to 25.5 on the Noometry Index.

Is Seed 2.0 Pro or Dolly 2.0-12b better for coding?

Seed 2.0 Pro scores higher on coding benchmarks: 43.5 versus 23.4 in the Noometry coding category.

How many benchmarks do Seed 2.0 Pro and Dolly 2.0-12b share?

9 benchmarks have published results for both models. Seed 2.0 Pro has 20 scored results on Noometry and Dolly 2.0-12b has 17.

Related comparisons

Go deeper