Model comparison

Qwen3.5 35B-A3B vs Qwen3.5-Flash

Qwen3.5 35B-A3B and Qwen3.5-Flash score almost the same on the Noometry Index (42.0 vs 42.5), so choose on price, context window or the category you care about most.

Last verified . 25 shared benchmarks.

Qwen3.5 35B-A3B Alibaba (Qwen)

42.0

Rank #123 Confirmed

Qwen3.5-Flash Alibaba (Qwen)

42.5

Rank #112 Confirmed

Summary

  • They share 25 benchmarks with published results for both. Qwen3.5 35B-A3B scores higher in 3 categories and Qwen3.5-Flash in 5 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Qwen3.5-Flash leads 33.7 to 24.6.
  • The biggest single-benchmark swing is OTIS Mock AIME 2024-2025: 70% for Qwen3.5 35B-A3B and 84.4% for Qwen3.5-Flash.
  • Qwen3.5-Flash is cheaper at $0.10 / $0.40 per million input/output tokens, against $0.25 / $2 for Qwen3.5 35B-A3B.
  • Qwen3.5-Flash accepts more context: 1M tokens versus 262K.
  • Qwen3.5 35B-A3B has downloadable open weights; the other is API-only.

Side by side

Qwen3.5 35B-A3B and Qwen3.5-Flash specifications
Qwen3.5 35B-A3BQwen3.5-Flash
ProviderAlibaba (Qwen)Alibaba (Qwen)
Noometry Index42.042.5
Released2026-02-012026-02-23
WeightsOpenProprietary
Context window262K1M
Max output66K66K
Input $ / M tokens$0.25$0.10
Output $ / M tokens$2$0.40
Results tracked2832

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Qwen3.5 35B-A3B: 33.8 (#251), Qwen3.5-Flash: 34.2 (#242)

Coding benchmarks
BenchmarkQwen3.5 35B-A3BQwen3.5-Flash
LMArena WebDev12541244
LMArena Coding14101412
SciCode29.3%—
ALE-Bench—221.8

Agentic & Tool Use Not comparable

Qwen3.5 35B-A3B: —, Qwen3.5-Flash: —

Agentic & Tool Use benchmarks
BenchmarkQwen3.5 35B-A3BQwen3.5-Flash
Vending-Bench 2—462.69

Reasoning Qwen3.5-Flash leads

Qwen3.5 35B-A3B: 24.6 (#161), Qwen3.5-Flash: 33.7 (#72)

Reasoning benchmarks
BenchmarkQwen3.5 35B-A3BQwen3.5-Flash
Chess Puzzles10%21%
LMArena Hard Prompts14001403
DTBench80%82.9%
LMCA29.5%29.1%
Epoch Capabilities Index142.52143.98
CritPt0.6%—
Mystery Game Puzzles—20%

Math Qwen3.5 35B-A3B leads

Qwen3.5 35B-A3B: 39.9 (#97), Qwen3.5-Flash: 37.4 (#158)

Knowledge Qwen3.5 35B-A3B leads

Qwen3.5 35B-A3B: 47.8 (#79), Qwen3.5-Flash: 43.2 (#93)

Knowledge benchmarks
BenchmarkQwen3.5 35B-A3BQwen3.5-Flash
GPQA Diamond83.5%82.3%
Vectara Hallucination Rate10.5%10.5%
LMArena Expert14081407
SimpleQA Verified—20.3%

Multilingual Too close to call

Qwen3.5 35B-A3B: 50.0 (#127), Qwen3.5-Flash: 50.5 (#121)

Multilingual benchmarks
BenchmarkQwen3.5 35B-A3BQwen3.5-Flash
LMArena Non-English13781385
LMArena Chinese14571446
LMArena French14121412
LMArena German13671390
LMArena Japanese13251368
LMArena Korean13561344
LMArena Russian13761379
LMArena Spanish13921400

Instruction Following Too close to call

Qwen3.5 35B-A3B: 72.8 (#128), Qwen3.5-Flash: 72.6 (#139)

Instruction Following benchmarks
BenchmarkQwen3.5 35B-A3BQwen3.5-Flash
LMArena Instruction Following13791374

Long Context Too close to call

Qwen3.5 35B-A3B: 42.4 (#127), Qwen3.5-Flash: 42.4 (#124)

Long Context benchmarks
BenchmarkQwen3.5 35B-A3BQwen3.5-Flash
LMArena Longer Query13891392

Writing & Preference Too close to call

Qwen3.5 35B-A3B: 57.9 (#124), Qwen3.5-Flash: 57.9 (#122)

Writing & Preference benchmarks
BenchmarkQwen3.5 35B-A3BQwen3.5-Flash
LMArena Text13951397
LMArena Creative Writing13461343
LMArena Multi-Turn13901393

Frequently asked questions

Is Qwen3.5 35B-A3B better than Qwen3.5-Flash?

Qwen3.5 35B-A3B and Qwen3.5-Flash score almost the same on the Noometry Index (42.0 vs 42.5), so choose on price, context window or the category you care about most.

Which is cheaper, Qwen3.5 35B-A3B or Qwen3.5-Flash?

Qwen3.5-Flash is cheaper. It lists at $0.10 per million input tokens and $0.40 per million output tokens; Qwen3.5 35B-A3B lists at $0.25 and $2.

Is Qwen3.5 35B-A3B or Qwen3.5-Flash better for coding?

They score almost the same on coding (33.8 vs 34.2); test both on your own repository before choosing.

Which has the bigger context window?

Qwen3.5-Flash does, with 1M tokens against 262K.

How many benchmarks do Qwen3.5 35B-A3B and Qwen3.5-Flash share?

25 benchmarks have published results for both models. Qwen3.5 35B-A3B has 28 scored results on Noometry and Qwen3.5-Flash has 32.

Related comparisons

Go deeper