Model comparison

Gemini 1.0 Pro vs Qwen3.6 Plus

Qwen3.6 Plus is the stronger model overall, scoring 47.5 to 27.3 on the Noometry Index.

Last verified . 20 shared benchmarks.

Gemini 1.0 Pro Google

27.3

Rank #332 Confirmed

Qwen3.6 Plus Alibaba (Qwen)

47.5

Rank #62 Confirmed

Summary

  • They share 20 benchmarks with published results for both. Gemini 1.0 Pro scores higher in 0 categories and Qwen3.6 Plus in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Qwen3.6 Plus leads 51.8 to 9.3.
  • The biggest single-benchmark swing is OTIS Mock AIME 2024-2025: 1.1% for Gemini 1.0 Pro and 93.3% for Qwen3.6 Plus.

Side by side

Gemini 1.0 Pro and Qwen3.6 Plus specifications
Gemini 1.0 ProQwen3.6 Plus
ProviderGoogleAlibaba (Qwen)
Noometry Index27.347.5
Released2023-12-132026-03-31
WeightsProprietaryProprietary
Context window—1M
Max output—66K
Input $ / M tokens—$0.50
Output $ / M tokens—$3
Results tracked2437

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.6 Plus leads

Gemini 1.0 Pro: 32.2 (#275), Qwen3.6 Plus: 40.8 (#130)

Coding benchmarks
BenchmarkGemini 1.0 ProQwen3.6 Plus
LMArena Coding11081467
SWE-bench Verified—57.9%
LMArena WebDev—1461
SciCode—40.7%
ALE-Bench—670.15
HumanEval+55.5%—
MBPP+61.4%—

Agentic & Tool Use Not comparable

Gemini 1.0 Pro: —, Qwen3.6 Plus: —

Agentic & Tool Use benchmarks
BenchmarkGemini 1.0 ProQwen3.6 Plus
Vending-Bench 2—5,115

Reasoning Qwen3.6 Plus leads

Gemini 1.0 Pro: 17.1 (#296), Qwen3.6 Plus: 29.3 (#93)

Reasoning benchmarks
BenchmarkGemini 1.0 ProQwen3.6 Plus
LMArena Hard Prompts11091449
DTBench45.9%81.9%
Epoch Capabilities Index117.04147.65
NYT Connections (extended)—60.3%
CritPt—2.9%
Chess Puzzles—17%
Thematic Generalization—59.5%
Mystery Game Puzzles—12%
LMCA—33.1%

Math Qwen3.6 Plus leads

Gemini 1.0 Pro: 9.3 (#321), Qwen3.6 Plus: 51.8 (#54)

Math benchmarks
BenchmarkGemini 1.0 ProQwen3.6 Plus
OTIS Mock AIME 2024-20251.1%93.3%
LMArena Math11321450
FrontierMath (Tiers 1-3)—38.2%
MATH Level 511.2%—
FrontierMath (Feb 2025 set)—26.2%
FrontierMath Tier 4 (v1)—8.3%

Knowledge Qwen3.6 Plus leads

Gemini 1.0 Pro: 15.6 (#291), Qwen3.6 Plus: 56.1 (#45)

Knowledge benchmarks
BenchmarkGemini 1.0 ProQwen3.6 Plus
GPQA Diamond34%88.4%
LMArena Expert10591454
SimpleQA Verified—44.1%
MMLU70%—

Multilingual Qwen3.6 Plus leads

Gemini 1.0 Pro: 33.4 (#252), Qwen3.6 Plus: 53.3 (#70)

Multilingual benchmarks
BenchmarkGemini 1.0 ProQwen3.6 Plus
LMArena Non-English11381424
LMArena Chinese11241477
LMArena French11451455
LMArena German11251452
LMArena Japanese10231389
LMArena Russian11861434
LMArena Spanish11191432
LMArena Korean—1379

Instruction Following Qwen3.6 Plus leads

Gemini 1.0 Pro: 57.6 (#267), Qwen3.6 Plus: 75.0 (#74)

Instruction Following benchmarks
BenchmarkGemini 1.0 ProQwen3.6 Plus
LMArena Instruction Following11141425

Long Context Qwen3.6 Plus leads

Gemini 1.0 Pro: 34.3 (#249), Qwen3.6 Plus: 45.2 (#49)

Long Context benchmarks
BenchmarkGemini 1.0 ProQwen3.6 Plus
LMArena Longer Query11321439
CL-bench—20.3%

Writing & Preference Qwen3.6 Plus leads

Gemini 1.0 Pro: 36.0 (#264), Qwen3.6 Plus: 62.2 (#82)

Writing & Preference benchmarks
BenchmarkGemini 1.0 ProQwen3.6 Plus
LMArena Text11491437
LMArena Creative Writing11311404
LMArena Multi-Turn11391438

Frequently asked questions

Is Gemini 1.0 Pro better than Qwen3.6 Plus?

Qwen3.6 Plus is the stronger model overall, scoring 47.5 to 27.3 on the Noometry Index.

Is Gemini 1.0 Pro or Qwen3.6 Plus better for coding?

Qwen3.6 Plus scores higher on coding benchmarks: 40.8 versus 32.2 in the Noometry coding category.

How many benchmarks do Gemini 1.0 Pro and Qwen3.6 Plus share?

20 benchmarks have published results for both models. Gemini 1.0 Pro has 24 scored results on Noometry and Qwen3.6 Plus has 37.

Related comparisons

Go deeper