Model comparison

Amazon Nova Experimental Chat 11 10 vs Gemini 3 Deep Think

Amazon Nova Experimental Chat 11 10 has enough public results to be ranked (#101); Gemini 3 Deep Think does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

Gemini 3 Deep Think Google

58.3

Unranked Sparse

Summary

  • The widest gap is in reasoning, where Gemini 3 Deep Think leads 70.2 to 29.1.

Side by side

Amazon Nova Experimental Chat 11 10 and Gemini 3 Deep Think specifications
Amazon Nova Experimental Chat 11 10Gemini 3 Deep Think
ProviderAmazonGoogle
Noometry Index43.058.3
Released—2026-02-12
WeightsProprietaryProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked173

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Amazon Nova Experimental Chat 11 10: 42.3 (#107), Gemini 3 Deep Think: —

Coding benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10Gemini 3 Deep Think
LMArena Coding1433—

Reasoning Gemini 3 Deep Think leads

Amazon Nova Experimental Chat 11 10: 29.1 (#95), Gemini 3 Deep Think: 70.2 (#14)

Reasoning benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10Gemini 3 Deep Think
ARC-AGI-2—84.6%
ARC-AGI-1—96%
CritPt—25.7%
LMArena Hard Prompts1422—

Math Not comparable

Amazon Nova Experimental Chat 11 10: 39.1 (#114), Gemini 3 Deep Think: —

Math benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10Gemini 3 Deep Think
LMArena Math1431—

Knowledge Not comparable

Amazon Nova Experimental Chat 11 10: 39.9 (#127), Gemini 3 Deep Think: —

Knowledge benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10Gemini 3 Deep Think
LMArena Expert1431—

Multilingual Not comparable

Amazon Nova Experimental Chat 11 10: 51.9 (#98), Gemini 3 Deep Think: —

Multilingual benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10Gemini 3 Deep Think
LMArena Non-English1405—
LMArena Chinese1463—
LMArena French1449—
LMArena German1395—
LMArena Japanese1383—
LMArena Korean1384—
LMArena Russian1394—
LMArena Spanish1444—

Instruction Following Not comparable

Amazon Nova Experimental Chat 11 10: 73.2 (#122), Gemini 3 Deep Think: —

Instruction Following benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10Gemini 3 Deep Think
LMArena Instruction Following1387—

Long Context Not comparable

Amazon Nova Experimental Chat 11 10: 42.5 (#121), Gemini 3 Deep Think: —

Long Context benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10Gemini 3 Deep Think
LMArena Longer Query1395—

Writing & Preference Not comparable

Amazon Nova Experimental Chat 11 10: 58.8 (#114), Gemini 3 Deep Think: —

Writing & Preference benchmarks
BenchmarkAmazon Nova Experimental Chat 11 10Gemini 3 Deep Think
LMArena Text1415—
LMArena Creative Writing1349—
LMArena Multi-Turn1380—

Frequently asked questions

Is Amazon Nova Experimental Chat 11 10 better than Gemini 3 Deep Think?

Amazon Nova Experimental Chat 11 10 has enough public results to be ranked (#101); Gemini 3 Deep Think does not yet, so treat this comparison as directional.

How many benchmarks do Amazon Nova Experimental Chat 11 10 and Gemini 3 Deep Think share?

0 benchmarks have published results for both models. Amazon Nova Experimental Chat 11 10 has 17 scored results on Noometry and Gemini 3 Deep Think has 3.

Related comparisons

Go deeper