Model comparison

Amazon Nova Lite vs o3

o3 is the stronger model overall, scoring 47.5 to 31.9 on the Noometry Index. Amazon Nova Lite costs 33× less per token, which makes it the better buy when o3's lead doesn't matter for your workload.

Last verified . 25 shared benchmarks.

Amazon Nova Lite Amazon

31.9

Rank #265 Confirmed

o3 OpenAI

47.5

Rank #61 Confirmed

Summary

  • They share 25 benchmarks with published results for both. Amazon Nova Lite scores higher in 0 categories and o3 in 9 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where o3 leads 54.6 to 26.3.
  • The biggest single-benchmark swing is Omni-MATH: 23.3% for Amazon Nova Lite and 71.4% for o3.
  • Amazon Nova Lite is cheaper at $0.06 / $0.24 per million input/output tokens, against $2 / $8 for o3.
  • Amazon Nova Lite accepts more context: 300K tokens versus 200K.

Side by side

Amazon Nova Lite and o3 specifications
Amazon Nova Liteo3
ProviderAmazonOpenAI
Noometry Index31.947.5
Released2024-12-032025-04-16
WeightsProprietaryProprietary
Context window300K200K
Max output10K100K
Input $ / M tokens$0.06$2
Output $ / M tokens$0.24$8
Results tracked3463

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding o3 leads

Amazon Nova Lite: 32.5 (#270), o3: 46.8 (#64)

Coding benchmarks
BenchmarkAmazon Nova Liteo3
LMArena Coding12391408
ALE-Bench236.25933.55
SWE-bench Verified—62.3%
SWE-bench Verified (bash only)—58.4%
Aider Polyglot—81.3%
GSO—8.8%
WeirdML—52.4%
LiveBench Coding27.5%—
CadEval—74%

Agentic & Tool Use Not comparable

Amazon Nova Lite: —, o3: 34.5 (#44)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova Liteo3
Berkeley Function Calling Leaderboard—63%
GDPval—30.8%
DeepResearch Bench—45.2%
OSWorld—23%
LMArena Search—1144
METR Time Horizons—65.4%

Reasoning o3 leads

Amazon Nova Lite: 19.3 (#260), o3: 32.0 (#78)

Reasoning benchmarks
BenchmarkAmazon Nova Liteo3
LMArena Hard Prompts12201402
ARC-AGI-2—6.5%
SimpleBench—53.1%
Kagi LLM Benchmark—67.6%
ARC-AGI-1—60.8%
CritPt—1.4%
Chess Puzzles—38%
EnigmaEval—13.1%
LiveBench Reasoning36.7%—
Mystery Game Puzzles—29%
DTBench—84.8%
LiveBench Data Analysis37.2%—
LMCA—39.7%
Epoch Capabilities Index—146.86
ForecastBench—62.5
LiveBench36.4%—

Math o3 leads

Amazon Nova Lite: 27.8 (#247), o3: 50.2 (#58)

Math benchmarks
BenchmarkAmazon Nova Liteo3
Omni-MATH23.3%71.4%
LMArena Math12271426
FrontierMath (Tiers 1-3)—33.3%
OTIS Mock AIME 2024-2025—84.4%
LiveBench Math36.7%—
MATH Level 5—97.8%
FrontierMath (Feb 2025 set)—18.7%
FrontierMath Tier 4 (v1)—2.1%

Knowledge o3 leads

Amazon Nova Lite: 26.3 (#259), o3: 54.6 (#52)

Knowledge benchmarks
BenchmarkAmazon Nova Liteo3
Humanity's Last Exam3.6%20.3%
MMLU-Pro60%85.9%
GPQA (HELM)39.7%75.3%
LMArena Expert12011402
GPQA Diamond—81.8%
SimpleQA Verified—49.4%
Confabulations—14.4%
Vectara Hallucination Rate6.1%—
MMLU77%—

Multimodal o3 leads

Amazon Nova Lite: 25.5 (#123), o3: 41.4 (#36)

Multimodal benchmarks
BenchmarkAmazon Nova Liteo3
LMArena Vision9901214
GeoBench—74%
VPCT—52%

Multilingual o3 leads

Amazon Nova Lite: 38.0 (#232), o3: 51.7 (#105)

Multilingual benchmarks
BenchmarkAmazon Nova Liteo3
LMArena Non-English12081401
LMArena Chinese12251437
LMArena French12371430
LMArena German12291420
LMArena Japanese11531403
LMArena Korean11531370
LMArena Russian12161406
LMArena Spanish12261395

Instruction Following o3 leads

Amazon Nova Lite: 59.3 (#255), o3: 72.8 (#127)

Instruction Following benchmarks
BenchmarkAmazon Nova Liteo3
IFEval77.6%86.9%
LMArena Instruction Following12051368
LiveBench Instruction Following54.1%—

Long Context o3 leads

Amazon Nova Lite: 37.4 (#217), o3: 53.3 (#6)

Long Context benchmarks
BenchmarkAmazon Nova Liteo3
LMArena Longer Query12341372
Fiction.LiveBench—88.9%
CL-bench—17.8%

Writing & Preference o3 leads

Amazon Nova Lite: 42.1 (#237), o3: 63.5 (#64)

Writing & Preference benchmarks
BenchmarkAmazon Nova Liteo3
LMArena Text12291410
LMArena Creative Writing11971359
WildBench75%86.1%
LMArena Multi-Turn11991405
Short-Story Creative Writing—83.9%
EQ-Bench Creative Writing—1676
LiveBench Language25.9%—

Frequently asked questions

Is Amazon Nova Lite better than o3?

o3 is the stronger model overall, scoring 47.5 to 31.9 on the Noometry Index. Amazon Nova Lite costs 33× less per token, which makes it the better buy when o3's lead doesn't matter for your workload.

Which is cheaper, Amazon Nova Lite or o3?

Amazon Nova Lite is cheaper. It lists at $0.06 per million input tokens and $0.24 per million output tokens; o3 lists at $2 and $8.

Is Amazon Nova Lite or o3 better for coding?

o3 scores higher on coding benchmarks: 46.8 versus 32.5 in the Noometry coding category.

Which has the bigger context window?

Amazon Nova Lite does, with 300K tokens against 200K.

How many benchmarks do Amazon Nova Lite and o3 share?

25 benchmarks have published results for both models. Amazon Nova Lite has 34 scored results on Noometry and o3 has 63.

Related comparisons

Go deeper