Model comparison

Llama 3.2 1B vs Nova 2.0 Pro Preview

Nova 2.0 Pro Preview is the stronger model overall, scoring 33.4 to 20.1 on the Noometry Index.

Last verified . 0 shared benchmarks.

Llama 3.2 1B Meta

20.1

Rank #354 Confirmed

Nova 2.0 Pro Preview Amazon

33.4

Rank #244 Reported

Summary

  • The widest gap is in coding, where Nova 2.0 Pro Preview leads 40.8 to 21.1.
  • Llama 3.2 1B has downloadable open weights; the other is API-only.

Side by side

Llama 3.2 1B and Nova 2.0 Pro Preview specifications
Llama 3.2 1BNova 2.0 Pro Preview
ProviderMetaAmazon
Noometry Index20.133.4
Released2024-09-242025-12-02
WeightsOpenProprietary
Context window60K—
Max output54K—
Input $ / M tokens$0.027—
Output $ / M tokens$0.20—
Results tracked223

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Nova 2.0 Pro Preview leads

Llama 3.2 1B: 21.1 (#338), Nova 2.0 Pro Preview: 40.8 (#133)

Coding benchmarks
BenchmarkLlama 3.2 1BNova 2.0 Pro Preview
SciCode—42.7%
BigCodeBench Instruct8.2%—
LMArena Coding1070—
BigCodeBench Complete11.3%—

Agentic & Tool Use Nova 2.0 Pro Preview leads

Llama 3.2 1B: 14.6 (#150), Nova 2.0 Pro Preview: 21.9 (#136)

Agentic & Tool Use benchmarks
BenchmarkLlama 3.2 1BNova 2.0 Pro Preview
Berkeley Function Calling Leaderboard10.8%—
BALROG6.6%—
GDP.pdf—2%

Reasoning Nova 2.0 Pro Preview leads

Llama 3.2 1B: 16.2 (#308), Nova 2.0 Pro Preview: 22.4 (#194)

Reasoning benchmarks
BenchmarkLlama 3.2 1BNova 2.0 Pro Preview
CritPt—0%
Chess Puzzles0%—
LMArena Hard Prompts1044—
Epoch Capabilities Index101.99—

Math Not comparable

Llama 3.2 1B: 10.4 (#313), Nova 2.0 Pro Preview: —

Math benchmarks
BenchmarkLlama 3.2 1BNova 2.0 Pro Preview
OTIS Mock AIME 2024-20250.6%—
LMArena Math1086—

Knowledge Not comparable

Llama 3.2 1B: 7.2 (#312), Nova 2.0 Pro Preview: —

Knowledge benchmarks
BenchmarkLlama 3.2 1BNova 2.0 Pro Preview
GPQA Diamond23.9%—
LMArena Expert1007—

Multilingual Not comparable

Llama 3.2 1B: 23.8 (#292), Nova 2.0 Pro Preview: —

Multilingual benchmarks
BenchmarkLlama 3.2 1BNova 2.0 Pro Preview
LMArena Non-English973—
LMArena Chinese959—
LMArena German1014—
LMArena Russian941—

Instruction Following Not comparable

Llama 3.2 1B: 52.4 (#290), Nova 2.0 Pro Preview: —

Instruction Following benchmarks
BenchmarkLlama 3.2 1BNova 2.0 Pro Preview
LMArena Instruction Following1031—

Long Context Not comparable

Llama 3.2 1B: 31.9 (#274), Nova 2.0 Pro Preview: —

Long Context benchmarks
BenchmarkLlama 3.2 1BNova 2.0 Pro Preview
LMArena Longer Query1050—

Writing & Preference Not comparable

Llama 3.2 1B: 21.3 (#310), Nova 2.0 Pro Preview: —

Writing & Preference benchmarks
BenchmarkLlama 3.2 1BNova 2.0 Pro Preview
LMArena Text1055—
LMArena Creative Writing1033—
EQ-Bench Creative Writing200—
LMArena Multi-Turn1030—

Frequently asked questions

Is Llama 3.2 1B better than Nova 2.0 Pro Preview?

Nova 2.0 Pro Preview is the stronger model overall, scoring 33.4 to 20.1 on the Noometry Index.

Is Llama 3.2 1B or Nova 2.0 Pro Preview better for coding?

Nova 2.0 Pro Preview scores higher on coding benchmarks: 40.8 versus 21.1 in the Noometry coding category.

How many benchmarks do Llama 3.2 1B and Nova 2.0 Pro Preview share?

0 benchmarks have published results for both models. Llama 3.2 1B has 22 scored results on Noometry and Nova 2.0 Pro Preview has 3.

Related comparisons

Go deeper