Model comparison

GPT-5 Nano vs Qwen Plus

Qwen Plus is the stronger model overall, scoring 37.1 to 33.5 on the Noometry Index. GPT-5 Nano costs 4.4× less per token, which makes it the better buy when Qwen Plus's lead doesn't matter for your workload.

Last verified . 20 shared benchmarks.

GPT-5 Nano OpenAI

33.5

Rank #241 Confirmed

Qwen Plus Alibaba (Qwen)

37.1

Rank #210 Confirmed

Summary

  • They share 20 benchmarks with published results for both. GPT-5 Nano scores higher in 4 categories and Qwen Plus in 4 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen Plus leads 52.2 to 39.1.
  • The biggest single-benchmark swing is OTIS Mock AIME 2024-2025: 81.1% for GPT-5 Nano and 17.8% for Qwen Plus.
  • GPT-5 Nano is cheaper at $0.05 / $0.40 per million input/output tokens, against $0.40 / $1.20 for Qwen Plus.
  • Qwen Plus accepts more context: 1M tokens versus 400K.

Side by side

GPT-5 Nano and Qwen Plus specifications
GPT-5 NanoQwen Plus
ProviderOpenAIAlibaba (Qwen)
Noometry Index33.537.1
Released2025-08-072024-01-25
WeightsProprietaryProprietary
Context window400K1M
Max output128K33K
Input $ / M tokens$0.05$0.40
Output $ / M tokens$0.40$1.20
Results tracked4920

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen Plus leads

GPT-5 Nano: 33.6 (#254), Qwen Plus: 38.9 (#167)

Coding benchmarks
BenchmarkGPT-5 NanoQwen Plus
LMArena Coding13511328
SWE-bench Verified (bash only)34.8%—
WeirdML38.1%—
ALE-Bench718.67—

Agentic & Tool Use Not comparable

GPT-5 Nano: 25.8 (#106), Qwen Plus: —

Agentic & Tool Use benchmarks
BenchmarkGPT-5 NanoQwen Plus
Terminal-Bench21.8%—
Berkeley Function Calling Leaderboard51.5%—

Reasoning Qwen Plus leads

GPT-5 Nano: 16.3 (#306), Qwen Plus: 28.4 (#107)

Reasoning benchmarks
BenchmarkGPT-5 NanoQwen Plus
Kagi LLM Benchmark62.2%63.3%
LMArena Hard Prompts13281317
DTBench62.7%81.1%
LMCA7.9%24%
ARC-AGI-22.6%—
ARC-AGI-120.7%—
Chess Puzzles27%—
Mystery Game Puzzles9%—
Epoch Capabilities Index139.38—
ForecastBench59.1—

Math GPT-5 Nano leads

GPT-5 Nano: 29.4 (#241), Qwen Plus: 23.3 (#271)

Math benchmarks
BenchmarkGPT-5 NanoQwen Plus
OTIS Mock AIME 2024-202581.1%17.8%
LMArena Math13171326
MATH Level 595.2%65.3%
FrontierMath (Feb 2025 set)8.3%1.7%
FrontierMath (Tiers 1-3)20%—
FrontierMath Tier 42.4%—
ProofBench12%—
Omni-MATH54.6%—
FrontierMath Tier 4 (v1)2.1%—

Knowledge GPT-5 Nano leads

GPT-5 Nano: 35.9 (#178), Qwen Plus: 27.4 (#251)

Knowledge benchmarks
BenchmarkGPT-5 NanoQwen Plus
GPQA Diamond69.4%48.1%
LMArena Expert13211328
SimpleQA Verified11.7%—
MMLU-Pro77.8%—
Vectara Hallucination Rate10.5%—
GPQA (HELM)67.9%—

Multimodal Not comparable

GPT-5 Nano: 31.3 (#108), Qwen Plus: —

Multimodal benchmarks
BenchmarkGPT-5 NanoQwen Plus
LMArena Vision1159—
VPCT37.2%—

Multilingual Too close to call

GPT-5 Nano: 45.3 (#172), Qwen Plus: 45.1 (#175)

Multilingual benchmarks
BenchmarkGPT-5 NanoQwen Plus
LMArena Non-English13131310
LMArena Chinese13561347
LMArena Japanese12261251
LMArena Russian12961323
LMArena German1327—
LMArena Korean1269—
LMArena Spanish1360—

Instruction Following GPT-5 Nano leads

GPT-5 Nano: 75.0 (#79), Qwen Plus: 68.8 (#181)

Instruction Following benchmarks
BenchmarkGPT-5 NanoQwen Plus
LMArena Instruction Following13061303
IFEval93.2%—

Long Context Qwen Plus leads

GPT-5 Nano: 31.3 (#281), Qwen Plus: 40.3 (#158)

Long Context benchmarks
BenchmarkGPT-5 NanoQwen Plus
LMArena Longer Query13121324
Fiction.LiveBench44.4%—

Writing & Preference Qwen Plus leads

GPT-5 Nano: 39.1 (#249), Qwen Plus: 52.2 (#176)

Writing & Preference benchmarks
BenchmarkGPT-5 NanoQwen Plus
LMArena Text13201326
LMArena Creative Writing12491293
LMArena Multi-Turn13111336
EQ-Bench Creative Writing705—
WildBench80.6%—

Frequently asked questions

Is GPT-5 Nano better than Qwen Plus?

Qwen Plus is the stronger model overall, scoring 37.1 to 33.5 on the Noometry Index. GPT-5 Nano costs 4.4× less per token, which makes it the better buy when Qwen Plus's lead doesn't matter for your workload.

Which is cheaper, GPT-5 Nano or Qwen Plus?

GPT-5 Nano is cheaper. It lists at $0.05 per million input tokens and $0.40 per million output tokens; Qwen Plus lists at $0.40 and $1.20.

Is GPT-5 Nano or Qwen Plus better for coding?

Qwen Plus scores higher on coding benchmarks: 38.9 versus 33.6 in the Noometry coding category.

Which has the bigger context window?

Qwen Plus does, with 1M tokens against 400K.

How many benchmarks do GPT-5 Nano and Qwen Plus share?

20 benchmarks have published results for both models. GPT-5 Nano has 49 scored results on Noometry and Qwen Plus has 20.

Related comparisons

Go deeper