Model comparison

Longcat Flash Chat vs Mistral Small 3.2

Longcat Flash Chat is the stronger model overall, scoring 42.1 to 31.2 on the Noometry Index.

Last verified . 1 shared benchmarks.

Longcat Flash Chat Meituan

42.1

Rank #120 Confirmed

Mistral Small 3.2 Mistral AI

31.2

Rank #280 Confirmed

Summary

  • They share 1 benchmark with published results for both. Longcat Flash Chat scores higher in 4 categories and Mistral Small 3.2 in 0 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Longcat Flash Chat leads 61.0 to 45.0.

Side by side

Longcat Flash Chat and Mistral Small 3.2 specifications
Longcat Flash ChatMistral Small 3.2
ProviderMeituanMistral AI
Noometry Index42.131.2
Released—2025-06-20
WeightsOpenOpen
Context window—256K
Max output—16K
Input $ / M tokens—$0.0938
Output $ / M tokens—$0.25
Results tracked196

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Longcat Flash Chat: 43.5 (#87), Mistral Small 3.2: —

Coding benchmarks
BenchmarkLongcat Flash ChatMistral Small 3.2
LMArena Coding1471—

Reasoning Too close to call

Longcat Flash Chat: 19.0 (#272), Mistral Small 3.2: 18.1 (#287)

Reasoning benchmarks
BenchmarkLongcat Flash ChatMistral Small 3.2
Kagi LLM Benchmark43.9%40.4%
NYT Connections (extended)17.7%—
Chess Puzzles—1%
LMArena Hard Prompts1440—
Epoch Capabilities Index—131.74

Math Longcat Flash Chat leads

Longcat Flash Chat: 39.4 (#107), Mistral Small 3.2: 26.3 (#260)

Math benchmarks
BenchmarkLongcat Flash ChatMistral Small 3.2
OTIS Mock AIME 2024-2025—30.3%
LMArena Math1442—

Knowledge Longcat Flash Chat leads

Longcat Flash Chat: 40.6 (#116), Mistral Small 3.2: 26.7 (#256)

Knowledge benchmarks
BenchmarkLongcat Flash ChatMistral Small 3.2
GPQA Diamond—49.1%
LMArena Expert1454—

Multilingual Not comparable

Longcat Flash Chat: 51.9 (#101), Mistral Small 3.2: —

Multilingual benchmarks
BenchmarkLongcat Flash ChatMistral Small 3.2
LMArena Non-English1404—
LMArena Chinese1465—
LMArena French1456—
LMArena German1408—
LMArena Japanese1373—
LMArena Korean1371—
LMArena Russian1395—
LMArena Spanish1445—

Instruction Following Not comparable

Longcat Flash Chat: 74.4 (#96), Mistral Small 3.2: —

Instruction Following benchmarks
BenchmarkLongcat Flash ChatMistral Small 3.2
LMArena Instruction Following1411—

Long Context Not comparable

Longcat Flash Chat: 43.5 (#93), Mistral Small 3.2: —

Long Context benchmarks
BenchmarkLongcat Flash ChatMistral Small 3.2
LMArena Longer Query1425—

Writing & Preference Longcat Flash Chat leads

Longcat Flash Chat: 61.0 (#91), Mistral Small 3.2: 45.0 (#224)

Writing & Preference benchmarks
BenchmarkLongcat Flash ChatMistral Small 3.2
LMArena Text1427—
LMArena Creative Writing1388—
EQ-Bench Creative Writing—1255
LMArena Multi-Turn1418—

Frequently asked questions

Is Longcat Flash Chat better than Mistral Small 3.2?

Longcat Flash Chat is the stronger model overall, scoring 42.1 to 31.2 on the Noometry Index.

How many benchmarks do Longcat Flash Chat and Mistral Small 3.2 share?

1 benchmark has published results for both models. Longcat Flash Chat has 19 scored results on Noometry and Mistral Small 3.2 has 6.

Related comparisons

Go deeper