Model comparison

Amazon Nova Micro vs DeepSeek V4.1 Flash

DeepSeek V4.1 Flash is the stronger model overall, scoring 52.8 to 30.4 on the Noometry Index. Amazon Nova Micro costs 4.3× less per token, which makes it the better buy when DeepSeek V4.1 Flash's lead doesn't matter for your workload.

Last verified . 17 shared benchmarks.

Amazon Nova Micro Amazon

30.4

Rank #294 Confirmed

DeepSeek V4.1 Flash DeepSeek

52.8

Rank #38 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Amazon Nova Micro scores higher in 0 categories and DeepSeek V4.1 Flash in 9 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in math, where DeepSeek V4.1 Flash leads 66.7 to 26.9.
  • Amazon Nova Micro is cheaper at $0.035 / $0.14 per million input/output tokens, against $0.15 / $0.60 for DeepSeek V4.1 Flash.
  • DeepSeek V4.1 Flash accepts more context: 1M tokens versus 128K.
  • DeepSeek V4.1 Flash has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Micro and DeepSeek V4.1 Flash specifications
Amazon Nova MicroDeepSeek V4.1 Flash
ProviderAmazonDeepSeek
Noometry Index30.452.8
Released2024-12-032026-09-09
WeightsProprietaryOpen
Context window128K1M
Max output10K393K
Input $ / M tokens$0.035$0.15
Output $ / M tokens$0.14$0.60
Results tracked3237

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding DeepSeek V4.1 Flash leads

Amazon Nova Micro: 30.5 (#295), DeepSeek V4.1 Flash: 52.9 (#32)

Coding benchmarks
BenchmarkAmazon Nova MicroDeepSeek V4.1 Flash
LMArena Coding12181506
LMArena WebDev—1619
SciCode—51.9%
LiveBench Coding20.2%—
ALE-Bench—1,092

Agentic & Tool Use DeepSeek V4.1 Flash leads

Amazon Nova Micro: 22.1 (#132), DeepSeek V4.1 Flash: 31.2 (#69)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova MicroDeepSeek V4.1 Flash
APEX-Agents—39.5%
Berkeley Function Calling Leaderboard22.3%—
GDP.pdf—19.8%

Reasoning DeepSeek V4.1 Flash leads

Amazon Nova Micro: 17.4 (#294), DeepSeek V4.1 Flash: 50.2 (#36)

Reasoning benchmarks
BenchmarkAmazon Nova MicroDeepSeek V4.1 Flash
LMArena Hard Prompts11911483
NYT Connections (extended)—89.6%
CritPt—14.3%
LiveBench Reasoning25.1%—
Mystery Game Puzzles—43%
DTBench—89.9%
LiveBench Data Analysis34%—
LMCA—47%
Surface Evolver Bench—46.3%
Epoch Capabilities Index—154.9
LiveBench29.6%—

Math DeepSeek V4.1 Flash leads

Amazon Nova Micro: 26.9 (#254), DeepSeek V4.1 Flash: 66.7 (#25)

Math benchmarks
BenchmarkAmazon Nova MicroDeepSeek V4.1 Flash
LMArena Math12061477
FrontierMath (Tiers 1-3)—67.4%
FrontierMath Tier 4—26.8%
OTIS Mock AIME 2024-2025—98.3%
ProofBench—54%
Omni-MATH21.4%—
LiveBench Math34.5%—

Knowledge DeepSeek V4.1 Flash leads

Amazon Nova Micro: 29.6 (#237), DeepSeek V4.1 Flash: 57.9 (#38)

Knowledge benchmarks
BenchmarkAmazon Nova MicroDeepSeek V4.1 Flash
LMArena Expert11841506
GPQA Diamond—89.8%
MMLU-Pro51.1%—
Vectara Hallucination Rate5.5%—
GPQA (HELM)38.3%—
MMLU70.8%—

Multimodal Not comparable

Amazon Nova Micro: —, DeepSeek V4.1 Flash: 39.1 (#61)

Multimodal benchmarks
BenchmarkAmazon Nova MicroDeepSeek V4.1 Flash
LMArena Vision—1277
Furniture Assembly—34.2%

Multilingual DeepSeek V4.1 Flash leads

Amazon Nova Micro: 36.5 (#239), DeepSeek V4.1 Flash: 55.0 (#35)

Multilingual benchmarks
BenchmarkAmazon Nova MicroDeepSeek V4.1 Flash
LMArena Non-English11861448
LMArena Chinese12091497
LMArena French12381452
LMArena German11921484
LMArena Japanese11541412
LMArena Korean11501452
LMArena Russian11851471
LMArena Spanish12251459

Instruction Following DeepSeek V4.1 Flash leads

Amazon Nova Micro: 56.3 (#272), DeepSeek V4.1 Flash: 77.3 (#26)

Instruction Following benchmarks
BenchmarkAmazon Nova MicroDeepSeek V4.1 Flash
LMArena Instruction Following11741474
LiveBench Instruction Following48%—
IFEval76%—

Long Context DeepSeek V4.1 Flash leads

Amazon Nova Micro: 36.5 (#229), DeepSeek V4.1 Flash: 45.2 (#47)

Long Context benchmarks
BenchmarkAmazon Nova MicroDeepSeek V4.1 Flash
LMArena Longer Query12051475

Writing & Preference DeepSeek V4.1 Flash leads

Amazon Nova Micro: 39.5 (#247), DeepSeek V4.1 Flash: 65.4 (#48)

Writing & Preference benchmarks
BenchmarkAmazon Nova MicroDeepSeek V4.1 Flash
LMArena Text12081462
LMArena Creative Writing11721435
LMArena Multi-Turn11781457
EQ-Bench Creative Writing—1540
WildBench74.3%—
LiveBench Language15.8%—

Frequently asked questions

Is Amazon Nova Micro better than DeepSeek V4.1 Flash?

DeepSeek V4.1 Flash is the stronger model overall, scoring 52.8 to 30.4 on the Noometry Index. Amazon Nova Micro costs 4.3× less per token, which makes it the better buy when DeepSeek V4.1 Flash's lead doesn't matter for your workload.

Which is cheaper, Amazon Nova Micro or DeepSeek V4.1 Flash?

Amazon Nova Micro is cheaper. It lists at $0.035 per million input tokens and $0.14 per million output tokens; DeepSeek V4.1 Flash lists at $0.15 and $0.60.

Is Amazon Nova Micro or DeepSeek V4.1 Flash better for coding?

DeepSeek V4.1 Flash scores higher on coding benchmarks: 52.9 versus 30.5 in the Noometry coding category.

Which has the bigger context window?

DeepSeek V4.1 Flash does, with 1M tokens against 128K.

How many benchmarks do Amazon Nova Micro and DeepSeek V4.1 Flash share?

17 benchmarks have published results for both models. Amazon Nova Micro has 32 scored results on Noometry and DeepSeek V4.1 Flash has 37.

Related comparisons

Go deeper