Model comparison

Longcat Flash Chat vs MiniMax-M2.7

Longcat Flash Chat is the stronger model overall, scoring 42.1 to 37.7 on the Noometry Index.

Last verified . 18 shared benchmarks.

Longcat Flash Chat Meituan

42.1

Rank #120 Confirmed

MiniMax-M2.7 MiniMax

37.7

Rank #196 Confirmed

Summary

  • They share 18 benchmarks with published results for both. Longcat Flash Chat scores higher in 7 categories and MiniMax-M2.7 in 1 category; 5 gaps are clear of the uncertainty.
  • The widest gap is in math, where Longcat Flash Chat leads 39.4 to 25.9.
  • The biggest single-benchmark swing is NYT Connections (extended): 17.7% for Longcat Flash Chat and 24.7% for MiniMax-M2.7.

Side by side

Longcat Flash Chat and MiniMax-M2.7 specifications
Longcat Flash ChatMiniMax-M2.7
ProviderMeituanMiniMax
Noometry Index42.137.7
Released—2026-03-18
WeightsOpenOpen
Context window—205K
Max output—131K
Input $ / M tokens—$0.30
Output $ / M tokens—$1.20
Results tracked1930

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Longcat Flash Chat leads

Longcat Flash Chat: 43.5 (#87), MiniMax-M2.7: 41.8 (#120)

Coding benchmarks
BenchmarkLongcat Flash ChatMiniMax-M2.7
LMArena Coding14711454
LMArena WebDev—1398
SciCode—47%
WeirdML—37%
ALE-Bench—599.25

Agentic & Tool Use Not comparable

Longcat Flash Chat: —, MiniMax-M2.7: 25.1 (#111)

Agentic & Tool Use benchmarks
BenchmarkLongcat Flash ChatMiniMax-M2.7
Terminal-Bench—45.1%
ExploitBench—13.3%
GBAEval—0%

Reasoning Too close to call

Longcat Flash Chat: 19.0 (#272), MiniMax-M2.7: 19.7 (#253)

Reasoning benchmarks
BenchmarkLongcat Flash ChatMiniMax-M2.7
NYT Connections (extended)17.7%24.7%
LMArena Hard Prompts14401422
Kagi LLM Benchmark43.9%—
CritPt—0.6%
Thematic Generalization—39.3%
Epoch Capabilities Index—145.85

Math Longcat Flash Chat leads

Longcat Flash Chat: 39.4 (#107), MiniMax-M2.7: 25.9 (#263)

Math benchmarks
BenchmarkLongcat Flash ChatMiniMax-M2.7
LMArena Math14421420
ProofBench—3%

Knowledge Longcat Flash Chat leads

Longcat Flash Chat: 40.6 (#116), MiniMax-M2.7: 37.7 (#152)

Knowledge benchmarks
BenchmarkLongcat Flash ChatMiniMax-M2.7
LMArena Expert14541444
Vectara Hallucination Rate—12.9%

Multilingual Longcat Flash Chat leads

Longcat Flash Chat: 51.9 (#101), MiniMax-M2.7: 50.3 (#123)

Multilingual benchmarks
BenchmarkLongcat Flash ChatMiniMax-M2.7
LMArena Non-English14041382
LMArena Chinese14651441
LMArena French14561421
LMArena German14081398
LMArena Japanese13731262
LMArena Korean13711313
LMArena Russian13951383
LMArena Spanish14451403

Instruction Following Too close to call

Longcat Flash Chat: 74.4 (#96), MiniMax-M2.7: 74.1 (#103)

Instruction Following benchmarks
BenchmarkLongcat Flash ChatMiniMax-M2.7
LMArena Instruction Following14111405

Long Context Too close to call

Longcat Flash Chat: 43.5 (#93), MiniMax-M2.7: 43.3 (#99)

Long Context benchmarks
BenchmarkLongcat Flash ChatMiniMax-M2.7
LMArena Longer Query14251419

Writing & Preference Longcat Flash Chat leads

Longcat Flash Chat: 61.0 (#91), MiniMax-M2.7: 58.9 (#112)

Writing & Preference benchmarks
BenchmarkLongcat Flash ChatMiniMax-M2.7
LMArena Text14271405
LMArena Creative Writing13881354
LMArena Multi-Turn14181412

Frequently asked questions

Is Longcat Flash Chat better than MiniMax-M2.7?

Longcat Flash Chat is the stronger model overall, scoring 42.1 to 37.7 on the Noometry Index.

Is Longcat Flash Chat or MiniMax-M2.7 better for coding?

Longcat Flash Chat scores higher on coding benchmarks: 43.5 versus 41.8 in the Noometry coding category.

How many benchmarks do Longcat Flash Chat and MiniMax-M2.7 share?

18 benchmarks have published results for both models. Longcat Flash Chat has 19 scored results on Noometry and MiniMax-M2.7 has 30.

Related comparisons

Go deeper