Model comparison

Olmo 3.1 32b Instruct vs Yi-Large

Olmo 3.1 32b Instruct has enough public results to be ranked (#168); Yi-Large does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

Yi-Large 01.AI

38.1

Unranked Sparse

Summary

  • Olmo 3.1 32b Instruct has downloadable open weights; the other is API-only.

Side by side

Olmo 3.1 32b Instruct and Yi-Large specifications
Olmo 3.1 32b InstructYi-Large
ProviderAllen Institute for AI (Ai2)01.AI
Noometry Index39.438.1
Released—2024-05-13
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked163

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Olmo 3.1 32b Instruct leads

Olmo 3.1 32b Instruct: 39.5 (#157), Yi-Large: 36.7 (#203)

Coding benchmarks
BenchmarkOlmo 3.1 32b InstructYi-Large
BigCodeBench Instruct—37.7%
LMArena Coding1347—
BigCodeBench Complete—47.2%

Reasoning Not comparable

Olmo 3.1 32b Instruct: 26.4 (#132), Yi-Large: —

Reasoning benchmarks
BenchmarkOlmo 3.1 32b InstructYi-Large
LMArena Hard Prompts1322—

Math Not comparable

Olmo 3.1 32b Instruct: 36.3 (#167), Yi-Large: —

Math benchmarks
BenchmarkOlmo 3.1 32b InstructYi-Large
LMArena Math1305—

Knowledge Not comparable

Olmo 3.1 32b Instruct: 36.1 (#175), Yi-Large: —

Knowledge benchmarks
BenchmarkOlmo 3.1 32b InstructYi-Large
LMArena Expert1308—
MMLU—79.3%

Multilingual Not comparable

Olmo 3.1 32b Instruct: 42.6 (#191), Yi-Large: —

Multilingual benchmarks
BenchmarkOlmo 3.1 32b InstructYi-Large
LMArena Non-English1275—
LMArena Chinese1304—
LMArena French1328—
LMArena German1282—
LMArena Korean1206—
LMArena Russian1268—
LMArena Spanish1336—

Instruction Following Not comparable

Olmo 3.1 32b Instruct: 68.6 (#187), Yi-Large: —

Instruction Following benchmarks
BenchmarkOlmo 3.1 32b InstructYi-Large
LMArena Instruction Following1299—

Long Context Not comparable

Olmo 3.1 32b Instruct: 39.9 (#166), Yi-Large: —

Long Context benchmarks
BenchmarkOlmo 3.1 32b InstructYi-Large
LMArena Longer Query1312—

Writing & Preference Not comparable

Olmo 3.1 32b Instruct: 50.2 (#185), Yi-Large: —

Writing & Preference benchmarks
BenchmarkOlmo 3.1 32b InstructYi-Large
LMArena Text1311—
LMArena Creative Writing1264—
LMArena Multi-Turn1309—

Frequently asked questions

Is Olmo 3.1 32b Instruct better than Yi-Large?

Olmo 3.1 32b Instruct has enough public results to be ranked (#168); Yi-Large does not yet, so treat this comparison as directional.

Is Olmo 3.1 32b Instruct or Yi-Large better for coding?

Olmo 3.1 32b Instruct scores higher on coding benchmarks: 39.5 versus 36.7 in the Noometry coding category.

How many benchmarks do Olmo 3.1 32b Instruct and Yi-Large share?

0 benchmarks have published results for both models. Olmo 3.1 32b Instruct has 16 scored results on Noometry and Yi-Large has 3.

Related comparisons

Go deeper