Model comparison

Gemini 3.5 Flash vs Qwen3.5 Max Preview

Gemini 3.5 Flash is the stronger model overall, scoring 54.2 to 45.3 on the Noometry Index.

Last verified . 17 shared benchmarks.

Gemini 3.5 Flash Google

54.2

Rank #32 Confirmed

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Gemini 3.5 Flash scores higher in 6 categories and Qwen3.5 Max Preview in 2 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Gemini 3.5 Flash leads 62.8 to 30.8.

Side by side

Gemini 3.5 Flash and Qwen3.5 Max Preview specifications
Gemini 3.5 FlashQwen3.5 Max Preview
ProviderGoogleAlibaba (Qwen)
Noometry Index54.245.3
Released2026-05-19—
WeightsProprietaryProprietary
Context window1.05M—
Max output66K—
Input $ / M tokens$1.50—
Output $ / M tokens$9—
Results tracked5417

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 3.5 Flash leads

Gemini 3.5 Flash: 49.4 (#49), Qwen3.5 Max Preview: 44.0 (#77)

Coding benchmarks
BenchmarkGemini 3.5 FlashQwen3.5 Max Preview
LMArena Coding14921487
SWE-bench Verified79.3%—
DeepSWE37.4%—
LMArena WebDev1499—
SciCode53.1%—
WeirdML62.6%—
ALE-Bench911.02—

Agentic & Tool Use Not comparable

Gemini 3.5 Flash: 24.7 (#114), Qwen3.5 Max Preview: —

Agentic & Tool Use benchmarks
BenchmarkGemini 3.5 FlashQwen3.5 Max Preview
APEX-Agents27.5%—
GBAEval6.7%—
GDP.pdf14%—
Vending-Bench 25,396—

Reasoning Gemini 3.5 Flash leads

Gemini 3.5 Flash: 62.8 (#18), Qwen3.5 Max Preview: 30.8 (#84)

Reasoning benchmarks
BenchmarkGemini 3.5 FlashQwen3.5 Max Preview
LMArena Hard Prompts14881483
ARC-AGI-272.1%—
SimpleBench76.7%—
NYT Connections (extended)92.6%—
ARC-AGI-192.5%—
CritPt13.1%—
Chess Puzzles50%—
EnigmaEval25.4%—
EBR-Bench4.8%—
Mystery Game Puzzles32%—
DTBench94.7%—
LMCA47.1%—
Surface Evolver Bench58.1%—
Epoch Capabilities Index154.46—
ForecastBench59—

Math Gemini 3.5 Flash leads

Gemini 3.5 Flash: 60.7 (#36), Qwen3.5 Max Preview: 40.1 (#94)

Math benchmarks
BenchmarkGemini 3.5 FlashQwen3.5 Max Preview
LMArena Math15041474
FrontierMath (Tiers 1-3)62.8%—
FrontierMath Tier 426.8%—
MathArena Final-Answer Competitions76.3%—
OTIS Mock AIME 2024-202595.6%—
ProofBench31%—
FrontierMath (Feb 2025 set)39%—
FrontierMath Tier 4 (v1)14.6%—

Knowledge Gemini 3.5 Flash leads

Gemini 3.5 Flash: 66.3 (#11), Qwen3.5 Max Preview: 41.8 (#107)

Knowledge benchmarks
BenchmarkGemini 3.5 FlashQwen3.5 Max Preview
LMArena Expert14951489
GPQA Diamond92.8%—
SimpleQA Verified66.2%—

Multimodal Not comparable

Gemini 3.5 Flash: 45.7 (#15), Qwen3.5 Max Preview: —

Multimodal benchmarks
BenchmarkGemini 3.5 FlashQwen3.5 Max Preview
LMArena Vision1310—
Blueprint-Bench 233.6%—
LMArena Document1463—

Multilingual Too close to call

Gemini 3.5 Flash: 57.0 (#13), Qwen3.5 Max Preview: 56.2 (#22)

Multilingual benchmarks
BenchmarkGemini 3.5 FlashQwen3.5 Max Preview
LMArena Non-English14761465
LMArena Chinese15261534
LMArena French14901484
LMArena German14921487
LMArena Japanese14861495
LMArena Korean14511438
LMArena Russian14931471
LMArena Spanish14801470

Instruction Following Too close to call

Gemini 3.5 Flash: 77.0 (#30), Qwen3.5 Max Preview: 77.0 (#31)

Instruction Following benchmarks
BenchmarkGemini 3.5 FlashQwen3.5 Max Preview
LMArena Instruction Following14671467

Long Context Too close to call

Gemini 3.5 Flash: 45.4 (#38), Qwen3.5 Max Preview: 45.2 (#45)

Long Context benchmarks
BenchmarkGemini 3.5 FlashQwen3.5 Max Preview
LMArena Longer Query14821476

Writing & Preference Too close to call

Gemini 3.5 Flash: 65.5 (#47), Qwen3.5 Max Preview: 66.0 (#41)

Writing & Preference benchmarks
BenchmarkGemini 3.5 FlashQwen3.5 Max Preview
LMArena Text14821470
LMArena Creative Writing14701464
LMArena Multi-Turn14811478
EQ-Bench 41087—

Frequently asked questions

Is Gemini 3.5 Flash better than Qwen3.5 Max Preview?

Gemini 3.5 Flash is the stronger model overall, scoring 54.2 to 45.3 on the Noometry Index.

Is Gemini 3.5 Flash or Qwen3.5 Max Preview better for coding?

Gemini 3.5 Flash scores higher on coding benchmarks: 49.4 versus 44.0 in the Noometry coding category.

How many benchmarks do Gemini 3.5 Flash and Qwen3.5 Max Preview share?

17 benchmarks have published results for both models. Gemini 3.5 Flash has 54 scored results on Noometry and Qwen3.5 Max Preview has 17.

Related comparisons

Go deeper