Model comparison

Claude 2 vs MiMo-V2-Omni

MiMo-V2-Omni is the stronger model overall, scoring 43.6 to 25.0 on the Noometry Index.

Last verified . 0 shared benchmarks.

Claude 2 Anthropic

25.0

Rank #346 Reported

MiMo-V2-Omni Xiaomi

43.6

Rank #88 Confirmed

Summary

  • The widest gap is in math, where MiMo-V2-Omni leads 39.1 to 9.3.

Side by side

Claude 2 and MiMo-V2-Omni specifications
Claude 2MiMo-V2-Omni
ProviderAnthropicXiaomi
Noometry Index25.043.6
Released2023-07-112026-03-18
WeightsProprietaryProprietary
Context window—262K
Max output—131K
Input $ / M tokens—$0.14
Output $ / M tokens—$0.28
Results tracked818

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Claude 2: —, MiMo-V2-Omni: 43.3 (#89)

Coding benchmarks
BenchmarkClaude 2MiMo-V2-Omni
LMArena Coding—1466
HumanEval+61.6%—

Reasoning MiMo-V2-Omni leads

Claude 2: 21.7 (#216), MiMo-V2-Omni: 29.7 (#88)

Reasoning benchmarks
BenchmarkClaude 2MiMo-V2-Omni
LMArena Hard Prompts—1445
DTBench51.9%—
Epoch Capabilities Index120.13—

Math MiMo-V2-Omni leads

Claude 2: 9.3 (#320), MiMo-V2-Omni: 39.1 (#115)

Math benchmarks
BenchmarkClaude 2MiMo-V2-Omni
OTIS Mock AIME 2024-20252.5%—
LMArena Math—1430
MATH Level 511.7%—

Knowledge MiMo-V2-Omni leads

Claude 2: 16.9 (#287), MiMo-V2-Omni: 40.5 (#118)

Knowledge benchmarks
BenchmarkClaude 2MiMo-V2-Omni
GPQA Diamond34.7%—
LMArena Expert—1449
MMLU78.5%—
TriviaQA87.5%—

Multimodal Not comparable

Claude 2: —, MiMo-V2-Omni: 38.6 (#63)

Multimodal benchmarks
BenchmarkClaude 2MiMo-V2-Omni
LMArena Vision—1228

Multilingual Not comparable

Claude 2: —, MiMo-V2-Omni: 51.8 (#102)

Multilingual benchmarks
BenchmarkClaude 2MiMo-V2-Omni
LMArena Non-English—1404
LMArena Chinese—1465
LMArena French—1447
LMArena German—1399
LMArena Japanese—1317
LMArena Korean—1355
LMArena Russian—1412
LMArena Spanish—1434

Instruction Following Not comparable

Claude 2: —, MiMo-V2-Omni: 75.2 (#66)

Instruction Following benchmarks
BenchmarkClaude 2MiMo-V2-Omni
LMArena Instruction Following—1428

Long Context Not comparable

Claude 2: —, MiMo-V2-Omni: 44.1 (#76)

Long Context benchmarks
BenchmarkClaude 2MiMo-V2-Omni
LMArena Longer Query—1442

Writing & Preference Not comparable

Claude 2: —, MiMo-V2-Omni: 61.4 (#87)

Writing & Preference benchmarks
BenchmarkClaude 2MiMo-V2-Omni
LMArena Text—1423
LMArena Creative Writing—1392
LMArena Multi-Turn—1445

Frequently asked questions

Is Claude 2 better than MiMo-V2-Omni?

MiMo-V2-Omni is the stronger model overall, scoring 43.6 to 25.0 on the Noometry Index.

How many benchmarks do Claude 2 and MiMo-V2-Omni share?

0 benchmarks have published results for both models. Claude 2 has 8 scored results on Noometry and MiMo-V2-Omni has 18.

Related comparisons

Go deeper