Model comparison

Grok Build 0.1 vs Nemotron 4 340b Instruct

Grok Build 0.1 and Nemotron 4 340b Instruct score almost the same on the Noometry Index (36.4 vs 35.9), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Grok Build 0.1 xAI

36.4

Rank #216 Reported

Nemotron 4 340b Instruct NVIDIA

35.9

Rank #223 Confirmed

Summary

  • The widest gap is in reasoning, where Grok Build 0.1 leads 32.2 to 23.4.
  • Nemotron 4 340b Instruct has downloadable open weights; the other is API-only.

Side by side

Grok Build 0.1 and Nemotron 4 340b Instruct specifications
Grok Build 0.1Nemotron 4 340b Instruct
ProviderxAINVIDIA
Noometry Index36.435.9
Released2026-04-162024-06-14
WeightsProprietaryOpen
Context window256K—
Max output256K—
Input $ / M tokens$1—
Output $ / M tokens$2—
Results tracked317

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Grok Build 0.1 leads

Grok Build 0.1: 43.1 (#91), Nemotron 4 340b Instruct: 35.2 (#228)

Coding benchmarks
BenchmarkGrok Build 0.1Nemotron 4 340b Instruct
SciCode50.2%—
LMArena Coding—1209

Agentic & Tool Use Not comparable

Grok Build 0.1: 22.7 (#129), Nemotron 4 340b Instruct: —

Agentic & Tool Use benchmarks
BenchmarkGrok Build 0.1Nemotron 4 340b Instruct
GBAEval2.4%—

Reasoning Grok Build 0.1 leads

Grok Build 0.1: 32.2 (#77), Nemotron 4 340b Instruct: 23.4 (#178)

Reasoning benchmarks
BenchmarkGrok Build 0.1Nemotron 4 340b Instruct
CritPt9.1%—
LMArena Hard Prompts—1202

Math Not comparable

Grok Build 0.1: —, Nemotron 4 340b Instruct: 34.3 (#195)

Math benchmarks
BenchmarkGrok Build 0.1Nemotron 4 340b Instruct
LMArena Math—1216

Knowledge Not comparable

Grok Build 0.1: —, Nemotron 4 340b Instruct: 32.0 (#216)

Knowledge benchmarks
BenchmarkGrok Build 0.1Nemotron 4 340b Instruct
LMArena Expert—1171

Multilingual Not comparable

Grok Build 0.1: —, Nemotron 4 340b Instruct: 37.9 (#233)

Multilingual benchmarks
BenchmarkGrok Build 0.1Nemotron 4 340b Instruct
LMArena Non-English—1206
LMArena Chinese—1214
LMArena French—1197
LMArena German—1192
LMArena Japanese—1128
LMArena Korean—1148
LMArena Russian—1219
LMArena Spanish—1195

Instruction Following Not comparable

Grok Build 0.1: —, Nemotron 4 340b Instruct: 63.1 (#232)

Instruction Following benchmarks
BenchmarkGrok Build 0.1Nemotron 4 340b Instruct
LMArena Instruction Following—1203

Long Context Not comparable

Grok Build 0.1: —, Nemotron 4 340b Instruct: 37.1 (#221)

Long Context benchmarks
BenchmarkGrok Build 0.1Nemotron 4 340b Instruct
LMArena Longer Query—1225

Writing & Preference Not comparable

Grok Build 0.1: —, Nemotron 4 340b Instruct: 42.7 (#233)

Writing & Preference benchmarks
BenchmarkGrok Build 0.1Nemotron 4 340b Instruct
LMArena Text—1225
LMArena Creative Writing—1203
LMArena Multi-Turn—1214

Frequently asked questions

Is Grok Build 0.1 better than Nemotron 4 340b Instruct?

Grok Build 0.1 and Nemotron 4 340b Instruct score almost the same on the Noometry Index (36.4 vs 35.9), so choose on price, context window or the category you care about most.

Is Grok Build 0.1 or Nemotron 4 340b Instruct better for coding?

Grok Build 0.1 scores higher on coding benchmarks: 43.1 versus 35.2 in the Noometry coding category.

How many benchmarks do Grok Build 0.1 and Nemotron 4 340b Instruct share?

0 benchmarks have published results for both models. Grok Build 0.1 has 3 scored results on Noometry and Nemotron 4 340b Instruct has 17.

Related comparisons

Go deeper