Model comparison

Amazon Nova Lite vs Qwen3 235B-A22B

Qwen3 235B-A22B is the stronger model overall, scoring 43.5 to 31.9 on the Noometry Index. Amazon Nova Lite costs 12× less per token, which makes it the better buy when Qwen3 235B-A22B's lead doesn't matter for your workload.

Last verified . 23 shared benchmarks.

Amazon Nova Lite Amazon

31.9

Rank #265 Confirmed

Qwen3 235B-A22B Alibaba (Qwen)

43.5

Rank #91 Confirmed

Summary

  • They share 23 benchmarks with published results for both. Amazon Nova Lite scores higher in 1 category and Qwen3 235B-A22B in 7 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Qwen3 235B-A22B leads 49.6 to 26.3.
  • The biggest single-benchmark swing is Omni-MATH: 23.3% for Amazon Nova Lite and 71.8% for Qwen3 235B-A22B.
  • Amazon Nova Lite is cheaper at $0.06 / $0.24 per million input/output tokens, against $0.70 / $2.80 for Qwen3 235B-A22B.
  • Amazon Nova Lite accepts more context: 300K tokens versus 131K.
  • Qwen3 235B-A22B has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Lite and Qwen3 235B-A22B specifications
Amazon Nova LiteQwen3 235B-A22B
ProviderAmazonAlibaba (Qwen)
Noometry Index31.943.5
Released2024-12-032025-04
WeightsProprietaryOpen
Context window300K131K
Max output10K16K
Input $ / M tokens$0.06$0.70
Output $ / M tokens$0.24$2.80
Results tracked3449

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3 235B-A22B leads

Amazon Nova Lite: 32.5 (#270), Qwen3 235B-A22B: 44.3 (#75)

Coding benchmarks
BenchmarkAmazon Nova LiteQwen3 235B-A22B
LMArena Coding12391445
Aider Polyglot—59.6%
SciCode—42.4%
WeirdML—41%
LiveBench Coding27.5%—
ALE-Bench236.25—

Agentic & Tool Use Not comparable

Amazon Nova Lite: —, Qwen3 235B-A22B: 33.9 (#51)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova LiteQwen3 235B-A22B
Berkeley Function Calling Leaderboard—52.1%
Vending-Bench 2—-11.34

Reasoning Amazon Nova Lite leads

Amazon Nova Lite: 19.3 (#260), Qwen3 235B-A22B: 15.7 (#311)

Reasoning benchmarks
BenchmarkAmazon Nova LiteQwen3 235B-A22B
LMArena Hard Prompts12201433
ARC-AGI-2—1.3%
SimpleBench—31%
Kagi LLM Benchmark—69.4%
ARC-AGI-1—11%
CritPt—0%
Chess Puzzles—12%
LiveBench Reasoning36.7%—
Mystery Game Puzzles—9%
DTBench—80.3%
LiveBench Data Analysis37.2%—
LMCA—29.3%
Epoch Capabilities Index—143.85
ForecastBench—59.7
LiveBench36.4%—

Math Qwen3 235B-A22B leads

Amazon Nova Lite: 27.8 (#247), Qwen3 235B-A22B: 50.4 (#57)

Math benchmarks
BenchmarkAmazon Nova LiteQwen3 235B-A22B
Omni-MATH23.3%71.8%
LMArena Math12271432
OTIS Mock AIME 2024-2025—86.7%
LiveBench Math36.7%—
MATH Level 5—68.9%
FrontierMath (Feb 2025 set)—8.5%
FrontierMath Tier 4 (v1)—0%

Knowledge Qwen3 235B-A22B leads

Amazon Nova Lite: 26.3 (#259), Qwen3 235B-A22B: 49.6 (#73)

Knowledge benchmarks
BenchmarkAmazon Nova LiteQwen3 235B-A22B
MMLU-Pro60%84.4%
Vectara Hallucination Rate6.1%9.3%
GPQA (HELM)39.7%72.7%
LMArena Expert12011463
GPQA Diamond—80.1%
Humanity's Last Exam3.6%—
SimpleQA Verified—40.4%
Confabulations—15.6%
MMLU77%—

Multimodal Not comparable

Amazon Nova Lite: 25.5 (#123), Qwen3 235B-A22B: —

Multimodal benchmarks
BenchmarkAmazon Nova LiteQwen3 235B-A22B
LMArena Vision990—

Multilingual Qwen3 235B-A22B leads

Amazon Nova Lite: 38.0 (#232), Qwen3 235B-A22B: 52.3 (#89)

Multilingual benchmarks
BenchmarkAmazon Nova LiteQwen3 235B-A22B
LMArena Non-English12081409
LMArena Chinese12251481
LMArena French12371445
LMArena German12291433
LMArena Japanese11531399
LMArena Korean11531391
LMArena Russian12161411
LMArena Spanish12261430

Instruction Following Qwen3 235B-A22B leads

Amazon Nova Lite: 59.3 (#255), Qwen3 235B-A22B: 72.6 (#136)

Instruction Following benchmarks
BenchmarkAmazon Nova LiteQwen3 235B-A22B
IFEval77.6%83.5%
LMArena Instruction Following12051408
LiveBench Instruction Following54.1%—

Long Context Qwen3 235B-A22B leads

Amazon Nova Lite: 37.4 (#217), Qwen3 235B-A22B: 46.1 (#26)

Long Context benchmarks
BenchmarkAmazon Nova LiteQwen3 235B-A22B
LMArena Longer Query12341426
Fiction.LiveBench—75%

Writing & Preference Qwen3 235B-A22B leads

Amazon Nova Lite: 42.1 (#237), Qwen3 235B-A22B: 59.6 (#108)

Writing & Preference benchmarks
BenchmarkAmazon Nova LiteQwen3 235B-A22B
LMArena Text12291419
LMArena Creative Writing11971384
WildBench75%86.6%
LMArena Multi-Turn11991432
Short-Story Creative Writing—83%
EQ-Bench Creative Writing—1366
LiveBench Language25.9%—

Frequently asked questions

Is Amazon Nova Lite better than Qwen3 235B-A22B?

Qwen3 235B-A22B is the stronger model overall, scoring 43.5 to 31.9 on the Noometry Index. Amazon Nova Lite costs 12× less per token, which makes it the better buy when Qwen3 235B-A22B's lead doesn't matter for your workload.

Which is cheaper, Amazon Nova Lite or Qwen3 235B-A22B?

Amazon Nova Lite is cheaper. It lists at $0.06 per million input tokens and $0.24 per million output tokens; Qwen3 235B-A22B lists at $0.70 and $2.80.

Is Amazon Nova Lite or Qwen3 235B-A22B better for coding?

Qwen3 235B-A22B scores higher on coding benchmarks: 44.3 versus 32.5 in the Noometry coding category.

Which has the bigger context window?

Amazon Nova Lite does, with 300K tokens against 131K.

How many benchmarks do Amazon Nova Lite and Qwen3 235B-A22B share?

23 benchmarks have published results for both models. Amazon Nova Lite has 34 scored results on Noometry and Qwen3 235B-A22B has 49.

Related comparisons

Go deeper