Model comparison

OLMo 2 Furious 13B vs Yi-Large

OLMo 2 Furious 13B has enough public results to be ranked (#304); Yi-Large does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

Yi-Large 01.AI

38.1

Unranked Sparse

Summary

  • The widest gap is in coding, where Yi-Large leads 36.7 to 28.6.
  • OLMo 2 Furious 13B has downloadable open weights; the other is API-only.

Side by side

OLMo 2 Furious 13B and Yi-Large specifications
OLMo 2 Furious 13BYi-Large
ProviderAllen Institute for AI (Ai2)01.AI
Noometry Index29.738.1
Released2024-12-312024-05-13
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked123

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Yi-Large leads

OLMo 2 Furious 13B: 28.6 (#313), Yi-Large: 36.7 (#203)

Coding benchmarks
BenchmarkOLMo 2 Furious 13BYi-Large
BigCodeBench Instruct—37.7%
LiveBench Coding10.4%—
BigCodeBench Complete—47.2%

Reasoning Not comparable

OLMo 2 Furious 13B: 15.5 (#314), Yi-Large: —

Reasoning benchmarks
BenchmarkOLMo 2 Furious 13BYi-Large
LiveBench Reasoning16.3%—
LiveBench Data Analysis20.6%—
LiveBench22.1%—

Math Not comparable

OLMo 2 Furious 13B: 23.1 (#273), Yi-Large: —

Math benchmarks
BenchmarkOLMo 2 Furious 13BYi-Large
Omni-MATH15.6%—
LiveBench Math13.6%—

Knowledge Not comparable

OLMo 2 Furious 13B: 18.3 (#283), Yi-Large: —

Knowledge benchmarks
BenchmarkOLMo 2 Furious 13BYi-Large
MMLU-Pro31%—
GPQA (HELM)31.6%—
MMLU—79.3%

Instruction Following Not comparable

OLMo 2 Furious 13B: 60.9 (#248), Yi-Large: —

Instruction Following benchmarks
BenchmarkOLMo 2 Furious 13BYi-Large
LiveBench Instruction Following60.6%—
IFEval73%—

Writing & Preference Not comparable

OLMo 2 Furious 13B: 42.9 (#230), Yi-Large: —

Writing & Preference benchmarks
BenchmarkOLMo 2 Furious 13BYi-Large
WildBench68.9%—
LiveBench Language11.2%—

Frequently asked questions

Is OLMo 2 Furious 13B better than Yi-Large?

OLMo 2 Furious 13B has enough public results to be ranked (#304); Yi-Large does not yet, so treat this comparison as directional.

Is OLMo 2 Furious 13B or Yi-Large better for coding?

Yi-Large scores higher on coding benchmarks: 36.7 versus 28.6 in the Noometry coding category.

How many benchmarks do OLMo 2 Furious 13B and Yi-Large share?

0 benchmarks have published results for both models. OLMo 2 Furious 13B has 12 scored results on Noometry and Yi-Large has 3.

Related comparisons

Go deeper