Model comparison

Amazon Nova Pro vs Grok Build 0.1

Grok Build 0.1 is the stronger model overall, scoring 36.4 to 31.0 on the Noometry Index.

Last verified . 0 shared benchmarks.

Amazon Nova Pro Amazon

31.0

Rank #281 Confirmed

Grok Build 0.1 xAI

36.4

Rank #216 Reported

Summary

  • The widest gap is in reasoning, where Grok Build 0.1 leads 32.2 to 20.0.
  • Grok Build 0.1 is cheaper at $1 / $2 per million input/output tokens, against $0.80 / $3.20 for Amazon Nova Pro.
  • Amazon Nova Pro accepts more context: 300K tokens versus 256K.

Side by side

Amazon Nova Pro and Grok Build 0.1 specifications
Amazon Nova ProGrok Build 0.1
ProviderAmazonxAI
Noometry Index31.036.4
Released2024-12-032026-04-16
WeightsProprietaryProprietary
Context window300K256K
Max output10K256K
Input $ / M tokens$0.80$1
Output $ / M tokens$3.20$2
Results tracked383

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Grok Build 0.1 leads

Amazon Nova Pro: 35.1 (#229), Grok Build 0.1: 43.1 (#91)

Coding benchmarks
BenchmarkAmazon Nova ProGrok Build 0.1
SciCode—50.2%
LiveBench Coding38.1%—
LMArena Coding1270—

Agentic & Tool Use Grok Build 0.1 leads

Amazon Nova Pro: 16.7 (#147), Grok Build 0.1: 22.7 (#129)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova ProGrok Build 0.1
Berkeley Function Calling Leaderboard25%—
TheAgentCompany1.7%—
GBAEval—2.4%

Reasoning Grok Build 0.1 leads

Amazon Nova Pro: 20.0 (#243), Grok Build 0.1: 32.2 (#77)

Reasoning benchmarks
BenchmarkAmazon Nova ProGrok Build 0.1
CritPt—9.1%
LiveBench Reasoning32.6%—
LMArena Hard Prompts1246—
LiveBench Data Analysis48.3%—
Epoch Capabilities Index123.8—
LiveBench43.5%—

Math Not comparable

Amazon Nova Pro: 28.5 (#243), Grok Build 0.1: —

Math benchmarks
BenchmarkAmazon Nova ProGrok Build 0.1
Omni-MATH24.2%—
LiveBench Math38%—
LMArena Math1252—

Knowledge Not comparable

Amazon Nova Pro: 27.4 (#250), Grok Build 0.1: —

Knowledge benchmarks
BenchmarkAmazon Nova ProGrok Build 0.1
Humanity's Last Exam4.4%—
MMLU-Pro67.3%—
Confabulations30.1%—
Vectara Hallucination Rate5.1%—
GPQA (HELM)44.6%—
LMArena Expert1211—
MMLU82%—

Multimodal Not comparable

Amazon Nova Pro: 25.0 (#126), Grok Build 0.1: —

Multimodal benchmarks
BenchmarkAmazon Nova ProGrok Build 0.1
LMArena Vision980—

Multilingual Not comparable

Amazon Nova Pro: 39.7 (#223), Grok Build 0.1: —

Multilingual benchmarks
BenchmarkAmazon Nova ProGrok Build 0.1
LMArena Non-English1234—
LMArena Chinese1244—
LMArena French1271—
LMArena German1243—
LMArena Japanese1200—
LMArena Korean1203—
LMArena Russian1240—
LMArena Spanish1182—

Instruction Following Not comparable

Amazon Nova Pro: 64.9 (#226), Grok Build 0.1: —

Instruction Following benchmarks
BenchmarkAmazon Nova ProGrok Build 0.1
LiveBench Instruction Following67.1%—
IFEval81.5%—
LMArena Instruction Following1235—

Long Context Not comparable

Amazon Nova Pro: 38.1 (#205), Grok Build 0.1: —

Long Context benchmarks
BenchmarkAmazon Nova ProGrok Build 0.1
LMArena Longer Query1255—

Writing & Preference Not comparable

Amazon Nova Pro: 43.9 (#226), Grok Build 0.1: —

Writing & Preference benchmarks
BenchmarkAmazon Nova ProGrok Build 0.1
LMArena Text1259—
LMArena Creative Writing1212—
Short-Story Creative Writing60.5%—
WildBench77.7%—
LMArena Multi-Turn1246—
LiveBench Language37%—

Frequently asked questions

Is Amazon Nova Pro better than Grok Build 0.1?

Grok Build 0.1 is the stronger model overall, scoring 36.4 to 31.0 on the Noometry Index.

Which is cheaper, Amazon Nova Pro or Grok Build 0.1?

Grok Build 0.1 is cheaper. It lists at $1 per million input tokens and $2 per million output tokens; Amazon Nova Pro lists at $0.80 and $3.20.

Is Amazon Nova Pro or Grok Build 0.1 better for coding?

Grok Build 0.1 scores higher on coding benchmarks: 43.1 versus 35.1 in the Noometry coding category.

Which has the bigger context window?

Amazon Nova Pro does, with 300K tokens against 256K.

How many benchmarks do Amazon Nova Pro and Grok Build 0.1 share?

0 benchmarks have published results for both models. Amazon Nova Pro has 38 scored results on Noometry and Grok Build 0.1 has 3.

Related comparisons

Go deeper