Model comparison

Kimi K2.7 Code vs Longcat Flash Chat

Kimi K2.7 Code is the stronger model overall, scoring 43.3 to 42.1 on the Noometry Index.

Last verified . 0 shared benchmarks.

Kimi K2.7 Code Moonshot AI

43.3

Rank #94 Confirmed

Longcat Flash Chat Meituan

42.1

Rank #120 Confirmed

Summary

  • The widest gap is in reasoning, where Kimi K2.7 Code leads 39.0 to 19.0.

Side by side

Kimi K2.7 Code and Longcat Flash Chat specifications
Kimi K2.7 CodeLongcat Flash Chat
ProviderMoonshot AIMeituan
Noometry Index43.342.1
Released2026-06-12—
WeightsOpenOpen
Context window262K—
Max output262K—
Input $ / M tokens$0.95—
Output $ / M tokens$4—
Results tracked1919

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Kimi K2.7 Code: 42.9 (#95), Longcat Flash Chat: 43.5 (#87)

Coding benchmarks
BenchmarkKimi K2.7 CodeLongcat Flash Chat
DeepSWE30.5%—
FrontierCode30.1%—
LMArena WebDev1473—
SciCode47.5%—
WeirdML54.1%—
LMArena Coding—1471
ALE-Bench886.23—

Agentic & Tool Use Not comparable

Kimi K2.7 Code: 24.0 (#122), Longcat Flash Chat: —

Agentic & Tool Use benchmarks
BenchmarkKimi K2.7 CodeLongcat Flash Chat
APEX-Agents37.6%—
GBAEval0.9%—
Vending-Bench 25,083—

Reasoning Kimi K2.7 Code leads

Kimi K2.7 Code: 39.0 (#61), Longcat Flash Chat: 19.0 (#272)

Reasoning benchmarks
BenchmarkKimi K2.7 CodeLongcat Flash Chat
SimpleBench57.9%—
Kagi LLM Benchmark—43.9%
NYT Connections (extended)—17.7%
CritPt10%—
Chess Puzzles21%—
LMArena Hard Prompts—1440
Surface Evolver Bench48.8%—
Epoch Capabilities Index149.97—

Math Kimi K2.7 Code leads

Kimi K2.7 Code: 52.9 (#48), Longcat Flash Chat: 39.4 (#107)

Math benchmarks
BenchmarkKimi K2.7 CodeLongcat Flash Chat
FrontierMath (Tiers 1-3)54%—
FrontierMath Tier 412.2%—
OTIS Mock AIME 2024-202595.6%—
LMArena Math—1442

Knowledge Kimi K2.7 Code leads

Kimi K2.7 Code: 53.5 (#57), Longcat Flash Chat: 40.6 (#116)

Knowledge benchmarks
BenchmarkKimi K2.7 CodeLongcat Flash Chat
GPQA Diamond87.9%—
SimpleQA Verified36.5%—
LMArena Expert—1454

Multilingual Not comparable

Kimi K2.7 Code: —, Longcat Flash Chat: 51.9 (#101)

Multilingual benchmarks
BenchmarkKimi K2.7 CodeLongcat Flash Chat
LMArena Non-English—1404
LMArena Chinese—1465
LMArena French—1456
LMArena German—1408
LMArena Japanese—1373
LMArena Korean—1371
LMArena Russian—1395
LMArena Spanish—1445

Instruction Following Not comparable

Kimi K2.7 Code: —, Longcat Flash Chat: 74.4 (#96)

Instruction Following benchmarks
BenchmarkKimi K2.7 CodeLongcat Flash Chat
LMArena Instruction Following—1411

Long Context Not comparable

Kimi K2.7 Code: —, Longcat Flash Chat: 43.5 (#93)

Long Context benchmarks
BenchmarkKimi K2.7 CodeLongcat Flash Chat
LMArena Longer Query—1425

Writing & Preference Not comparable

Kimi K2.7 Code: —, Longcat Flash Chat: 61.0 (#91)

Writing & Preference benchmarks
BenchmarkKimi K2.7 CodeLongcat Flash Chat
LMArena Text—1427
LMArena Creative Writing—1388
LMArena Multi-Turn—1418

Frequently asked questions

Is Kimi K2.7 Code better than Longcat Flash Chat?

Kimi K2.7 Code is the stronger model overall, scoring 43.3 to 42.1 on the Noometry Index.

Is Kimi K2.7 Code or Longcat Flash Chat better for coding?

They score almost the same on coding (42.9 vs 43.5); test both on your own repository before choosing.

How many benchmarks do Kimi K2.7 Code and Longcat Flash Chat share?

0 benchmarks have published results for both models. Kimi K2.7 Code has 19 scored results on Noometry and Longcat Flash Chat has 19.

Related comparisons

Go deeper