Model comparison

Amazon Nova Micro vs Llama 3.2 90B

Amazon Nova Micro is the stronger model overall, scoring 30.4 to 27.5 on the Noometry Index.

Last verified . 1 shared benchmarks.

Amazon Nova Micro Amazon

30.4

Rank #294 Confirmed

Llama 3.2 90B Meta

27.5

Rank #331 Confirmed

Summary

  • They share 1 benchmark with published results for both. Amazon Nova Micro scores higher in 2 categories and Llama 3.2 90B in 2 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in math, where Amazon Nova Micro leads 26.9 to 11.1.
  • Llama 3.2 90B has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Micro and Llama 3.2 90B specifications
Amazon Nova MicroLlama 3.2 90B
ProviderAmazonMeta
Noometry Index30.427.5
Released2024-12-032024-09-24
WeightsProprietaryOpen
Context window128K—
Max output10K—
Input $ / M tokens$0.035—
Output $ / M tokens$0.14—
Results tracked329

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Amazon Nova Micro: 30.5 (#295), Llama 3.2 90B: —

Coding benchmarks
BenchmarkAmazon Nova MicroLlama 3.2 90B
LiveBench Coding20.2%—
LMArena Coding1218—

Agentic & Tool Use Llama 3.2 90B leads

Amazon Nova Micro: 22.1 (#132), Llama 3.2 90B: 30.0 (#80)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova MicroLlama 3.2 90B
Berkeley Function Calling Leaderboard22.3%—
BALROG—27.3%

Reasoning Llama 3.2 90B leads

Amazon Nova Micro: 17.4 (#294), Llama 3.2 90B: 21.7 (#217)

Reasoning benchmarks
BenchmarkAmazon Nova MicroLlama 3.2 90B
EnigmaEval—0.4%
LiveBench Reasoning25.1%—
LMArena Hard Prompts1191—
LiveBench Data Analysis34%—
Epoch Capabilities Index—125.5
LiveBench29.6%—

Math Amazon Nova Micro leads

Amazon Nova Micro: 26.9 (#254), Llama 3.2 90B: 11.1 (#308)

Math benchmarks
BenchmarkAmazon Nova MicroLlama 3.2 90B
OTIS Mock AIME 2024-2025—2.6%
Omni-MATH21.4%—
LiveBench Math34.5%—
LMArena Math1206—
MATH Level 5—39.4%

Knowledge Amazon Nova Micro leads

Amazon Nova Micro: 29.6 (#237), Llama 3.2 90B: 21.7 (#274)

Knowledge benchmarks
BenchmarkAmazon Nova MicroLlama 3.2 90B
MMLU70.8%80.3%
GPQA Diamond—41%
MMLU-Pro51.1%—
Vectara Hallucination Rate5.5%—
GPQA (HELM)38.3%—
LMArena Expert1184—

Multimodal Not comparable

Amazon Nova Micro: —, Llama 3.2 90B: 25.4 (#124)

Multimodal benchmarks
BenchmarkAmazon Nova MicroLlama 3.2 90B
LMArena Vision—1000
GeoBench—52%

Multilingual Not comparable

Amazon Nova Micro: 36.5 (#239), Llama 3.2 90B: —

Multilingual benchmarks
BenchmarkAmazon Nova MicroLlama 3.2 90B
LMArena Non-English1186—
LMArena Chinese1209—
LMArena French1238—
LMArena German1192—
LMArena Japanese1154—
LMArena Korean1150—
LMArena Russian1185—
LMArena Spanish1225—

Instruction Following Not comparable

Amazon Nova Micro: 56.3 (#272), Llama 3.2 90B: —

Instruction Following benchmarks
BenchmarkAmazon Nova MicroLlama 3.2 90B
LiveBench Instruction Following48%—
IFEval76%—
LMArena Instruction Following1174—

Long Context Not comparable

Amazon Nova Micro: 36.5 (#229), Llama 3.2 90B: —

Long Context benchmarks
BenchmarkAmazon Nova MicroLlama 3.2 90B
LMArena Longer Query1205—

Writing & Preference Not comparable

Amazon Nova Micro: 39.5 (#247), Llama 3.2 90B: —

Writing & Preference benchmarks
BenchmarkAmazon Nova MicroLlama 3.2 90B
LMArena Text1208—
LMArena Creative Writing1172—
WildBench74.3%—
LMArena Multi-Turn1178—
LiveBench Language15.8%—

Frequently asked questions

Is Amazon Nova Micro better than Llama 3.2 90B?

Amazon Nova Micro is the stronger model overall, scoring 30.4 to 27.5 on the Noometry Index.

How many benchmarks do Amazon Nova Micro and Llama 3.2 90B share?

1 benchmark has published results for both models. Amazon Nova Micro has 32 scored results on Noometry and Llama 3.2 90B has 9.

Related comparisons

Go deeper