Model comparison

Gemini 1.0 Pro vs Qwen3.5 Plus

Qwen3.5 Plus is the stronger model overall, scoring 42.9 to 27.3 on the Noometry Index.

Last verified . 4 shared benchmarks.

Gemini 1.0 Pro Google

27.3

Rank #332 Confirmed

Qwen3.5 Plus Alibaba (Qwen)

42.9

Rank #106 Confirmed

Summary

  • They share 4 benchmarks with published results for both. Gemini 1.0 Pro scores higher in 0 categories and Qwen3.5 Plus in 4 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in math, where Qwen3.5 Plus leads 49.6 to 9.3.
  • The biggest single-benchmark swing is OTIS Mock AIME 2024-2025: 1.1% for Gemini 1.0 Pro and 86.7% for Qwen3.5 Plus.

Side by side

Gemini 1.0 Pro and Qwen3.5 Plus specifications
Gemini 1.0 ProQwen3.5 Plus
ProviderGoogleAlibaba (Qwen)
Noometry Index27.342.9
Released2023-12-132026-02-16
WeightsProprietaryProprietary
Context window—1M
Max output—66K
Input $ / M tokens—$0.40
Output $ / M tokens—$2.40
Results tracked2415

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Gemini 1.0 Pro: 32.2 (#275), Qwen3.5 Plus: —

Coding benchmarks
BenchmarkGemini 1.0 ProQwen3.5 Plus
LMArena Coding1108—
ALE-Bench—621.92
HumanEval+55.5%—
MBPP+61.4%—

Agentic & Tool Use Not comparable

Gemini 1.0 Pro: —, Qwen3.5 Plus: —

Agentic & Tool Use benchmarks
BenchmarkGemini 1.0 ProQwen3.5 Plus
Vending-Bench 2—0.54

Reasoning Qwen3.5 Plus leads

Gemini 1.0 Pro: 17.1 (#296), Qwen3.5 Plus: 32.8 (#74)

Reasoning benchmarks
BenchmarkGemini 1.0 ProQwen3.5 Plus
DTBench45.9%80.5%
Epoch Capabilities Index117.04146.78
Chess Puzzles—22%
LMArena Hard Prompts1109—
Mystery Game Puzzles—17%
LMCA—36.4%

Math Qwen3.5 Plus leads

Gemini 1.0 Pro: 9.3 (#321), Qwen3.5 Plus: 49.6 (#61)

Math benchmarks
BenchmarkGemini 1.0 ProQwen3.5 Plus
OTIS Mock AIME 2024-20251.1%86.7%
LMArena Math1132—
MATH Level 511.2%—
FrontierMath (Feb 2025 set)—21%
FrontierMath Tier 4 (v1)—2.1%

Knowledge Qwen3.5 Plus leads

Gemini 1.0 Pro: 15.6 (#291), Qwen3.5 Plus: 46.0 (#83)

Knowledge benchmarks
BenchmarkGemini 1.0 ProQwen3.5 Plus
GPQA Diamond34%84.8%
SimpleQA Verified—25.4%
Vectara Hallucination Rate—10.7%
LMArena Expert1059—
MMLU70%—

Multilingual Not comparable

Gemini 1.0 Pro: 33.4 (#252), Qwen3.5 Plus: —

Multilingual benchmarks
BenchmarkGemini 1.0 ProQwen3.5 Plus
LMArena Non-English1138—
LMArena Chinese1124—
LMArena French1145—
LMArena German1125—
LMArena Japanese1023—
LMArena Russian1186—
LMArena Spanish1119—

Instruction Following Not comparable

Gemini 1.0 Pro: 57.6 (#267), Qwen3.5 Plus: —

Instruction Following benchmarks
BenchmarkGemini 1.0 ProQwen3.5 Plus
LMArena Instruction Following1114—

Long Context Qwen3.5 Plus leads

Gemini 1.0 Pro: 34.3 (#249), Qwen3.5 Plus: 43.0 (#113)

Long Context benchmarks
BenchmarkGemini 1.0 ProQwen3.5 Plus
CL-bench—19.8%
CL-bench Life—12.4%
LMArena Longer Query1132—

Writing & Preference Not comparable

Gemini 1.0 Pro: 36.0 (#264), Qwen3.5 Plus: —

Writing & Preference benchmarks
BenchmarkGemini 1.0 ProQwen3.5 Plus
LMArena Text1149—
LMArena Creative Writing1131—
LMArena Multi-Turn1139—

Frequently asked questions

Is Gemini 1.0 Pro better than Qwen3.5 Plus?

Qwen3.5 Plus is the stronger model overall, scoring 42.9 to 27.3 on the Noometry Index.

How many benchmarks do Gemini 1.0 Pro and Qwen3.5 Plus share?

4 benchmarks have published results for both models. Gemini 1.0 Pro has 24 scored results on Noometry and Qwen3.5 Plus has 15.

Related comparisons

Go deeper