Model comparison

Olmo 2 0325 32b Instruct vs Yi-Large

Olmo 2 0325 32b Instruct has enough public results to be ranked (#254); Yi-Large does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

Yi-Large 01.AI

38.1

Unranked Sparse

Summary

  • Olmo 2 0325 32b Instruct has downloadable open weights; the other is API-only.

Side by side

Olmo 2 0325 32b Instruct and Yi-Large specifications
Olmo 2 0325 32b InstructYi-Large
ProviderAllen Institute for AI (Ai2)01.AI
Noometry Index32.738.1
Released—2024-05-13
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked163

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Yi-Large leads

Olmo 2 0325 32b Instruct: 35.2 (#227), Yi-Large: 36.7 (#203)

Coding benchmarks
BenchmarkOlmo 2 0325 32b InstructYi-Large
BigCodeBench Instruct—37.7%
LMArena Coding1210—
BigCodeBench Complete—47.2%

Reasoning Not comparable

Olmo 2 0325 32b Instruct: 23.6 (#175), Yi-Large: —

Reasoning benchmarks
BenchmarkOlmo 2 0325 32b InstructYi-Large
LMArena Hard Prompts1208—

Math Not comparable

Olmo 2 0325 32b Instruct: 26.8 (#255), Yi-Large: —

Math benchmarks
BenchmarkOlmo 2 0325 32b InstructYi-Large
Omni-MATH16.1%—
LMArena Math1208—

Knowledge Not comparable

Olmo 2 0325 32b Instruct: 19.5 (#279), Yi-Large: —

Knowledge benchmarks
BenchmarkOlmo 2 0325 32b InstructYi-Large
MMLU-Pro41.4%—
GPQA (HELM)28.7%—
MMLU—79.3%

Multilingual Not comparable

Olmo 2 0325 32b Instruct: 34.8 (#248), Yi-Large: —

Multilingual benchmarks
BenchmarkOlmo 2 0325 32b InstructYi-Large
LMArena Non-English1160—
LMArena Chinese1192—
LMArena Russian1187—

Instruction Following Not comparable

Olmo 2 0325 32b Instruct: 61.5 (#244), Yi-Large: —

Instruction Following benchmarks
BenchmarkOlmo 2 0325 32b InstructYi-Large
IFEval78%—
LMArena Instruction Following1186—

Long Context Not comparable

Olmo 2 0325 32b Instruct: 36.2 (#234), Yi-Large: —

Long Context benchmarks
BenchmarkOlmo 2 0325 32b InstructYi-Large
LMArena Longer Query1194—

Writing & Preference Not comparable

Olmo 2 0325 32b Instruct: 42.1 (#236), Yi-Large: —

Writing & Preference benchmarks
BenchmarkOlmo 2 0325 32b InstructYi-Large
LMArena Text1218—
LMArena Creative Writing1199—
WildBench73.4%—
LMArena Multi-Turn1221—

Frequently asked questions

Is Olmo 2 0325 32b Instruct better than Yi-Large?

Olmo 2 0325 32b Instruct has enough public results to be ranked (#254); Yi-Large does not yet, so treat this comparison as directional.

Is Olmo 2 0325 32b Instruct or Yi-Large better for coding?

Yi-Large scores higher on coding benchmarks: 36.7 versus 35.2 in the Noometry coding category.

How many benchmarks do Olmo 2 0325 32b Instruct and Yi-Large share?

0 benchmarks have published results for both models. Olmo 2 0325 32b Instruct has 16 scored results on Noometry and Yi-Large has 3.

Related comparisons

Go deeper