Model comparison

Kimi K2.7 Code vs Qwen3.5 Max Preview

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 43.3 on the Noometry Index.

Last verified . 0 shared benchmarks.

Kimi K2.7 Code Moonshot AI

43.3

Rank #94 Confirmed

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Summary

  • The widest gap is in math, where Kimi K2.7 Code leads 52.9 to 40.1.
  • Kimi K2.7 Code has downloadable open weights; the other is API-only.

Side by side

Kimi K2.7 Code and Qwen3.5 Max Preview specifications
Kimi K2.7 CodeQwen3.5 Max Preview
ProviderMoonshot AIAlibaba (Qwen)
Noometry Index43.345.3
Released2026-06-12—
WeightsOpenProprietary
Context window262K—
Max output262K—
Input $ / M tokens$0.95—
Output $ / M tokens$4—
Results tracked1917

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.5 Max Preview leads

Kimi K2.7 Code: 42.9 (#95), Qwen3.5 Max Preview: 44.0 (#77)

Coding benchmarks
BenchmarkKimi K2.7 CodeQwen3.5 Max Preview
DeepSWE30.5%—
FrontierCode30.1%—
LMArena WebDev1473—
SciCode47.5%—
WeirdML54.1%—
LMArena Coding—1487
ALE-Bench886.23—

Agentic & Tool Use Not comparable

Kimi K2.7 Code: 24.0 (#122), Qwen3.5 Max Preview: —

Agentic & Tool Use benchmarks
BenchmarkKimi K2.7 CodeQwen3.5 Max Preview
APEX-Agents37.6%—
GBAEval0.9%—
Vending-Bench 25,083—

Reasoning Kimi K2.7 Code leads

Kimi K2.7 Code: 39.0 (#61), Qwen3.5 Max Preview: 30.8 (#84)

Reasoning benchmarks
BenchmarkKimi K2.7 CodeQwen3.5 Max Preview
SimpleBench57.9%—
CritPt10%—
Chess Puzzles21%—
LMArena Hard Prompts—1483
Surface Evolver Bench48.8%—
Epoch Capabilities Index149.97—

Math Kimi K2.7 Code leads

Kimi K2.7 Code: 52.9 (#48), Qwen3.5 Max Preview: 40.1 (#94)

Math benchmarks
BenchmarkKimi K2.7 CodeQwen3.5 Max Preview
FrontierMath (Tiers 1-3)54%—
FrontierMath Tier 412.2%—
OTIS Mock AIME 2024-202595.6%—
LMArena Math—1474

Knowledge Kimi K2.7 Code leads

Kimi K2.7 Code: 53.5 (#57), Qwen3.5 Max Preview: 41.8 (#107)

Knowledge benchmarks
BenchmarkKimi K2.7 CodeQwen3.5 Max Preview
GPQA Diamond87.9%—
SimpleQA Verified36.5%—
LMArena Expert—1489

Multilingual Not comparable

Kimi K2.7 Code: —, Qwen3.5 Max Preview: 56.2 (#22)

Multilingual benchmarks
BenchmarkKimi K2.7 CodeQwen3.5 Max Preview
LMArena Non-English—1465
LMArena Chinese—1534
LMArena French—1484
LMArena German—1487
LMArena Japanese—1495
LMArena Korean—1438
LMArena Russian—1471
LMArena Spanish—1470

Instruction Following Not comparable

Kimi K2.7 Code: —, Qwen3.5 Max Preview: 77.0 (#31)

Instruction Following benchmarks
BenchmarkKimi K2.7 CodeQwen3.5 Max Preview
LMArena Instruction Following—1467

Long Context Not comparable

Kimi K2.7 Code: —, Qwen3.5 Max Preview: 45.2 (#45)

Long Context benchmarks
BenchmarkKimi K2.7 CodeQwen3.5 Max Preview
LMArena Longer Query—1476

Writing & Preference Not comparable

Kimi K2.7 Code: —, Qwen3.5 Max Preview: 66.0 (#41)

Writing & Preference benchmarks
BenchmarkKimi K2.7 CodeQwen3.5 Max Preview
LMArena Text—1470
LMArena Creative Writing—1464
LMArena Multi-Turn—1478

Frequently asked questions

Is Kimi K2.7 Code better than Qwen3.5 Max Preview?

Qwen3.5 Max Preview is the stronger model overall, scoring 45.3 to 43.3 on the Noometry Index.

Is Kimi K2.7 Code or Qwen3.5 Max Preview better for coding?

Qwen3.5 Max Preview scores higher on coding benchmarks: 44.0 versus 42.9 in the Noometry coding category.

How many benchmarks do Kimi K2.7 Code and Qwen3.5 Max Preview share?

0 benchmarks have published results for both models. Kimi K2.7 Code has 19 scored results on Noometry and Qwen3.5 Max Preview has 17.

Related comparisons

Go deeper