Model comparison

Gemini 1.0 Pro vs Longcat Flash Chat

Longcat Flash Chat is the stronger model overall, scoring 42.1 to 27.3 on the Noometry Index.

Last verified . 16 shared benchmarks.

Gemini 1.0 Pro Google

27.3

Rank #332 Confirmed

Longcat Flash Chat Meituan

42.1

Rank #120 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Gemini 1.0 Pro scores higher in 0 categories and Longcat Flash Chat in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Longcat Flash Chat leads 39.4 to 9.3.
  • Longcat Flash Chat has downloadable open weights; the other is API-only.

Side by side

Gemini 1.0 Pro and Longcat Flash Chat specifications
Gemini 1.0 ProLongcat Flash Chat
ProviderGoogleMeituan
Noometry Index27.342.1
Released2023-12-13—
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked2419

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Longcat Flash Chat leads

Gemini 1.0 Pro: 32.2 (#275), Longcat Flash Chat: 43.5 (#87)

Coding benchmarks
BenchmarkGemini 1.0 ProLongcat Flash Chat
LMArena Coding11081471
HumanEval+55.5%—
MBPP+61.4%—

Reasoning Longcat Flash Chat leads

Gemini 1.0 Pro: 17.1 (#296), Longcat Flash Chat: 19.0 (#272)

Reasoning benchmarks
BenchmarkGemini 1.0 ProLongcat Flash Chat
LMArena Hard Prompts11091440
Kagi LLM Benchmark—43.9%
NYT Connections (extended)—17.7%
DTBench45.9%—
Epoch Capabilities Index117.04—

Math Longcat Flash Chat leads

Gemini 1.0 Pro: 9.3 (#321), Longcat Flash Chat: 39.4 (#107)

Math benchmarks
BenchmarkGemini 1.0 ProLongcat Flash Chat
LMArena Math11321442
OTIS Mock AIME 2024-20251.1%—
MATH Level 511.2%—

Knowledge Longcat Flash Chat leads

Gemini 1.0 Pro: 15.6 (#291), Longcat Flash Chat: 40.6 (#116)

Knowledge benchmarks
BenchmarkGemini 1.0 ProLongcat Flash Chat
LMArena Expert10591454
GPQA Diamond34%—
MMLU70%—

Multilingual Longcat Flash Chat leads

Gemini 1.0 Pro: 33.4 (#252), Longcat Flash Chat: 51.9 (#101)

Multilingual benchmarks
BenchmarkGemini 1.0 ProLongcat Flash Chat
LMArena Non-English11381404
LMArena Chinese11241465
LMArena French11451456
LMArena German11251408
LMArena Japanese10231373
LMArena Russian11861395
LMArena Spanish11191445
LMArena Korean—1371

Instruction Following Longcat Flash Chat leads

Gemini 1.0 Pro: 57.6 (#267), Longcat Flash Chat: 74.4 (#96)

Instruction Following benchmarks
BenchmarkGemini 1.0 ProLongcat Flash Chat
LMArena Instruction Following11141411

Long Context Longcat Flash Chat leads

Gemini 1.0 Pro: 34.3 (#249), Longcat Flash Chat: 43.5 (#93)

Long Context benchmarks
BenchmarkGemini 1.0 ProLongcat Flash Chat
LMArena Longer Query11321425

Writing & Preference Longcat Flash Chat leads

Gemini 1.0 Pro: 36.0 (#264), Longcat Flash Chat: 61.0 (#91)

Writing & Preference benchmarks
BenchmarkGemini 1.0 ProLongcat Flash Chat
LMArena Text11491427
LMArena Creative Writing11311388
LMArena Multi-Turn11391418

Frequently asked questions

Is Gemini 1.0 Pro better than Longcat Flash Chat?

Longcat Flash Chat is the stronger model overall, scoring 42.1 to 27.3 on the Noometry Index.

Is Gemini 1.0 Pro or Longcat Flash Chat better for coding?

Longcat Flash Chat scores higher on coding benchmarks: 43.5 versus 32.2 in the Noometry coding category.

How many benchmarks do Gemini 1.0 Pro and Longcat Flash Chat share?

16 benchmarks have published results for both models. Gemini 1.0 Pro has 24 scored results on Noometry and Longcat Flash Chat has 19.

Related comparisons

Go deeper