Model comparison

Kimi K2.5 Instant vs Mistral Large 3

Kimi K2.5 Instant is the stronger model overall, scoring 43.6 to 39.1 on the Noometry Index.

Last verified . 18 shared benchmarks.

Kimi K2.5 Instant Moonshot AI

43.6

Rank #89 Confirmed

Mistral Large 3 Mistral AI

39.1

Rank #176 Confirmed

Summary

  • They share 18 benchmarks with published results for both. Kimi K2.5 Instant scores higher in 8 categories and Mistral Large 3 in 1 category; 5 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Kimi K2.5 Instant leads 29.7 to 15.2.

Side by side

Kimi K2.5 Instant and Mistral Large 3 specifications
Kimi K2.5 InstantMistral Large 3
ProviderMoonshot AIMistral AI
Noometry Index43.639.1
Released—2025-12-02
WeightsOpenOpen
Context window—262K
Max output—8K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.75
Results tracked1824

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Kimi K2.5 Instant leads

Kimi K2.5 Instant: 42.6 (#97), Mistral Large 3: 34.4 (#237)

Coding benchmarks
BenchmarkKimi K2.5 InstantMistral Large 3
LMArena WebDev14041230
LMArena Coding14841448

Reasoning Kimi K2.5 Instant leads

Kimi K2.5 Instant: 29.7 (#90), Mistral Large 3: 15.2 (#319)

Reasoning benchmarks
BenchmarkKimi K2.5 InstantMistral Large 3
LMArena Hard Prompts14431429
Kagi LLM Benchmark—50.9%
NYT Connections (extended)—7.5%
Thematic Generalization—23%

Math Too close to call

Kimi K2.5 Instant: 39.4 (#105), Mistral Large 3: 38.7 (#129)

Math benchmarks
BenchmarkKimi K2.5 InstantMistral Large 3
LMArena Math14421414

Knowledge Kimi K2.5 Instant leads

Kimi K2.5 Instant: 40.2 (#123), Mistral Large 3: 36.0 (#177)

Knowledge benchmarks
BenchmarkKimi K2.5 InstantMistral Large 3
LMArena Expert14401421
Vectara Hallucination Rate—14.5%

Multimodal Kimi K2.5 Instant leads

Kimi K2.5 Instant: 40.2 (#50), Mistral Large 3: 38.2 (#66)

Multimodal benchmarks
BenchmarkKimi K2.5 InstantMistral Large 3
LMArena Vision12541221

Multilingual Too close to call

Kimi K2.5 Instant: 52.0 (#94), Mistral Large 3: 52.5 (#84)

Multilingual benchmarks
BenchmarkKimi K2.5 InstantMistral Large 3
LMArena Non-English14061413
LMArena Chinese14491447
LMArena French14031455
LMArena German14131437
LMArena Korean13781384
LMArena Russian14041411
LMArena Spanish14471440
LMArena Japanese—1394

Instruction Following Kimi K2.5 Instant leads

Kimi K2.5 Instant: 75.3 (#65), Mistral Large 3: 74.0 (#108)

Instruction Following benchmarks
BenchmarkKimi K2.5 InstantMistral Large 3
LMArena Instruction Following14301403

Long Context Too close to call

Kimi K2.5 Instant: 43.9 (#83), Mistral Large 3: 43.1 (#105)

Long Context benchmarks
BenchmarkKimi K2.5 InstantMistral Large 3
LMArena Longer Query14351413

Writing & Preference Too close to call

Kimi K2.5 Instant: 60.6 (#95), Mistral Large 3: 60.0 (#101)

Writing & Preference benchmarks
BenchmarkKimi K2.5 InstantMistral Large 3
LMArena Text14201428
LMArena Creative Writing13811386
LMArena Multi-Turn14271429
EQ-Bench Creative Writing—1412

Frequently asked questions

Is Kimi K2.5 Instant better than Mistral Large 3?

Kimi K2.5 Instant is the stronger model overall, scoring 43.6 to 39.1 on the Noometry Index.

Is Kimi K2.5 Instant or Mistral Large 3 better for coding?

Kimi K2.5 Instant scores higher on coding benchmarks: 42.6 versus 34.4 in the Noometry coding category.

How many benchmarks do Kimi K2.5 Instant and Mistral Large 3 share?

18 benchmarks have published results for both models. Kimi K2.5 Instant has 18 scored results on Noometry and Mistral Large 3 has 24.

Related comparisons

Go deeper