Model comparison

Llama 3.1 Tulu 3 8b vs Nemotron 3 Nano 30B A3B

Nemotron 3 Nano 30B A3B is the stronger model overall, scoring 40.6 to 35.7 on the Noometry Index.

Last verified . 11 shared benchmarks.

Summary

  • They share 11 benchmarks with published results for both. Llama 3.1 Tulu 3 8b scores higher in 0 categories and Nemotron 3 Nano 30B A3B in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Nemotron 3 Nano 30B A3B leads 52.3 to 39.7.

Side by side

Llama 3.1 Tulu 3 8b and Nemotron 3 Nano 30B A3B specifications
Llama 3.1 Tulu 3 8bNemotron 3 Nano 30B A3B
ProviderAllen Institute for AI (Ai2)NVIDIA
Noometry Index35.740.6
Released—2025-12-15
WeightsOpenOpen
Context window—262K
Max output—262K
Input $ / M tokens—$0.05
Output $ / M tokens—$0.20
Results tracked1118

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Nemotron 3 Nano 30B A3B leads

Llama 3.1 Tulu 3 8b: 34.4 (#235), Nemotron 3 Nano 30B A3B: 40.5 (#138)

Coding benchmarks
BenchmarkLlama 3.1 Tulu 3 8bNemotron 3 Nano 30B A3B
LMArena Coding11831378

Reasoning Nemotron 3 Nano 30B A3B leads

Llama 3.1 Tulu 3 8b: 22.8 (#188), Nemotron 3 Nano 30B A3B: 27.1 (#124)

Reasoning benchmarks
BenchmarkLlama 3.1 Tulu 3 8bNemotron 3 Nano 30B A3B
LMArena Hard Prompts11741348

Math Nemotron 3 Nano 30B A3B leads

Llama 3.1 Tulu 3 8b: 33.9 (#198), Nemotron 3 Nano 30B A3B: 37.8 (#147)

Math benchmarks
BenchmarkLlama 3.1 Tulu 3 8bNemotron 3 Nano 30B A3B
LMArena Math11951371

Knowledge Not comparable

Llama 3.1 Tulu 3 8b: —, Nemotron 3 Nano 30B A3B: 38.2 (#148)

Knowledge benchmarks
BenchmarkLlama 3.1 Tulu 3 8bNemotron 3 Nano 30B A3B
Vectara Hallucination Rate—9.6%
LMArena Expert—1355

Multilingual Nemotron 3 Nano 30B A3B leads

Llama 3.1 Tulu 3 8b: 35.4 (#246), Nemotron 3 Nano 30B A3B: 45.8 (#162)

Multilingual benchmarks
BenchmarkLlama 3.1 Tulu 3 8bNemotron 3 Nano 30B A3B
LMArena Non-English11691320
LMArena Chinese11761382
LMArena Russian11931303
LMArena French—1396
LMArena German—1348
LMArena Japanese—1268
LMArena Korean—1269
LMArena Spanish—1376

Instruction Following Nemotron 3 Nano 30B A3B leads

Llama 3.1 Tulu 3 8b: 61.3 (#246), Nemotron 3 Nano 30B A3B: 68.8 (#180)

Instruction Following benchmarks
BenchmarkLlama 3.1 Tulu 3 8bNemotron 3 Nano 30B A3B
LMArena Instruction Following11741304

Long Context Nemotron 3 Nano 30B A3B leads

Llama 3.1 Tulu 3 8b: 35.8 (#239), Nemotron 3 Nano 30B A3B: 39.5 (#175)

Long Context benchmarks
BenchmarkLlama 3.1 Tulu 3 8bNemotron 3 Nano 30B A3B
LMArena Longer Query11811300

Writing & Preference Nemotron 3 Nano 30B A3B leads

Llama 3.1 Tulu 3 8b: 39.7 (#245), Nemotron 3 Nano 30B A3B: 52.3 (#174)

Writing & Preference benchmarks
BenchmarkLlama 3.1 Tulu 3 8bNemotron 3 Nano 30B A3B
LMArena Text11931348
LMArena Creative Writing11821269
LMArena Multi-Turn11541319

Frequently asked questions

Is Llama 3.1 Tulu 3 8b better than Nemotron 3 Nano 30B A3B?

Nemotron 3 Nano 30B A3B is the stronger model overall, scoring 40.6 to 35.7 on the Noometry Index.

Is Llama 3.1 Tulu 3 8b or Nemotron 3 Nano 30B A3B better for coding?

Nemotron 3 Nano 30B A3B scores higher on coding benchmarks: 40.5 versus 34.4 in the Noometry coding category.

How many benchmarks do Llama 3.1 Tulu 3 8b and Nemotron 3 Nano 30B A3B share?

11 benchmarks have published results for both models. Llama 3.1 Tulu 3 8b has 11 scored results on Noometry and Nemotron 3 Nano 30B A3B has 18.

Related comparisons

Go deeper