Model comparison

Grok Build 0.1 vs Yi-Lightning

Grok Build 0.1 and Yi-Lightning score almost the same on the Noometry Index (36.4 vs 37.1), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Grok Build 0.1 xAI

36.4

Rank #216 Reported

Yi-Lightning 01.AI

37.1

Rank #209 Confirmed

Summary

  • The widest gap is in coding, where Grok Build 0.1 leads 43.1 to 27.4.

Side by side

Grok Build 0.1 and Yi-Lightning specifications
Grok Build 0.1Yi-Lightning
ProviderxAI01.AI
Noometry Index36.437.1
Released2026-04-162024-12-02
WeightsProprietaryProprietary
Context window256K—
Max output256K—
Input $ / M tokens$1—
Output $ / M tokens$2—
Results tracked318

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Grok Build 0.1 leads

Grok Build 0.1: 43.1 (#91), Yi-Lightning: 27.4 (#320)

Coding benchmarks
BenchmarkGrok Build 0.1Yi-Lightning
Aider Polyglot—12.9%
SciCode50.2%—
LMArena Coding—1312

Agentic & Tool Use Not comparable

Grok Build 0.1: 22.7 (#129), Yi-Lightning: —

Agentic & Tool Use benchmarks
BenchmarkGrok Build 0.1Yi-Lightning
GBAEval2.4%—

Reasoning Grok Build 0.1 leads

Grok Build 0.1: 32.2 (#77), Yi-Lightning: 25.9 (#139)

Reasoning benchmarks
BenchmarkGrok Build 0.1Yi-Lightning
CritPt9.1%—
LMArena Hard Prompts—1302

Math Not comparable

Grok Build 0.1: —, Yi-Lightning: 36.2 (#172)

Math benchmarks
BenchmarkGrok Build 0.1Yi-Lightning
LMArena Math—1300

Knowledge Not comparable

Grok Build 0.1: —, Yi-Lightning: 35.4 (#185)

Knowledge benchmarks
BenchmarkGrok Build 0.1Yi-Lightning
LMArena Expert—1286

Multilingual Not comparable

Grok Build 0.1: —, Yi-Lightning: 42.2 (#196)

Multilingual benchmarks
BenchmarkGrok Build 0.1Yi-Lightning
LMArena Non-English—1269
LMArena Chinese—1320
LMArena French—1305
LMArena German—1267
LMArena Japanese—1228
LMArena Korean—1192
LMArena Russian—1255
LMArena Spanish—1315

Instruction Following Not comparable

Grok Build 0.1: —, Yi-Lightning: 67.4 (#196)

Instruction Following benchmarks
BenchmarkGrok Build 0.1Yi-Lightning
LMArena Instruction Following—1278

Long Context Not comparable

Grok Build 0.1: —, Yi-Lightning: 39.4 (#181)

Long Context benchmarks
BenchmarkGrok Build 0.1Yi-Lightning
LMArena Longer Query—1297

Writing & Preference Not comparable

Grok Build 0.1: —, Yi-Lightning: 50.2 (#184)

Writing & Preference benchmarks
BenchmarkGrok Build 0.1Yi-Lightning
LMArena Text—1302
LMArena Creative Writing—1280
LMArena Multi-Turn—1311

Frequently asked questions

Is Grok Build 0.1 better than Yi-Lightning?

Grok Build 0.1 and Yi-Lightning score almost the same on the Noometry Index (36.4 vs 37.1), so choose on price, context window or the category you care about most.

Is Grok Build 0.1 or Yi-Lightning better for coding?

Grok Build 0.1 scores higher on coding benchmarks: 43.1 versus 27.4 in the Noometry coding category.

How many benchmarks do Grok Build 0.1 and Yi-Lightning share?

0 benchmarks have published results for both models. Grok Build 0.1 has 3 scored results on Noometry and Yi-Lightning has 18.

Related comparisons

Go deeper