Model comparison

GPT-6 Luna vs MiMo-V2-Omni

GPT-6 Luna is the stronger model overall, scoring 53.3 to 43.6 on the Noometry Index.

Last verified . 18 shared benchmarks.

GPT-6 Luna OpenAI

53.3

Rank #36 Confirmed

MiMo-V2-Omni Xiaomi

43.6

Rank #88 Confirmed

Summary

  • They share 18 benchmarks with published results for both. GPT-6 Luna scores higher in 5 categories and MiMo-V2-Omni in 4 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where GPT-6 Luna leads 76.1 to 39.1.
  • MiMo-V2-Omni is cheaper at $0.14 / $0.28 per million input/output tokens, against $0.10 / $0.50 for GPT-6 Luna.
  • GPT-6 Luna accepts more context: 1.05M tokens versus 262K.

Side by side

GPT-6 Luna and MiMo-V2-Omni specifications
GPT-6 LunaMiMo-V2-Omni
ProviderOpenAIXiaomi
Noometry Index53.343.6
Released2026-09-222026-03-18
WeightsProprietaryProprietary
Context window1.05M262K
Max output128K131K
Input $ / M tokens$0.10$0.14
Output $ / M tokens$0.50$0.28
Results tracked4218

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GPT-6 Luna leads

GPT-6 Luna: 55.5 (#25), MiMo-V2-Omni: 43.3 (#89)

Coding benchmarks
BenchmarkGPT-6 LunaMiMo-V2-Omni
LMArena Coding14391466
DeepSWE66.6%—
FrontierCode42.4%—
LMArena WebDev1581—
SciCode54.6%—
ALE-Bench1,577—

Agentic & Tool Use Not comparable

GPT-6 Luna: 33.3 (#54), MiMo-V2-Omni: —

Agentic & Tool Use benchmarks
BenchmarkGPT-6 LunaMiMo-V2-Omni
APEX-Agents44.3%—
GDP.pdf23%—

Reasoning GPT-6 Luna leads

GPT-6 Luna: 48.2 (#41), MiMo-V2-Omni: 29.7 (#88)

Reasoning benchmarks
BenchmarkGPT-6 LunaMiMo-V2-Omni
LMArena Hard Prompts14111445
ARC-AGI-259.3%—
NYT Connections (extended)68.7%—
ARC-AGI-186.7%—
CritPt19.4%—
Chess Puzzles31%—
Mystery Game Puzzles7%—
DTBench90.1%—
LMCA44.5%—
Epoch Capabilities Index156.28—

Math GPT-6 Luna leads

GPT-6 Luna: 76.1 (#15), MiMo-V2-Omni: 39.1 (#115)

Math benchmarks
BenchmarkGPT-6 LunaMiMo-V2-Omni
LMArena Math14161430
FrontierMath (Tiers 1-3)78.9%—
FrontierMath Tier 456.1%—
OTIS Mock AIME 2024-202598.9%—
ProofBench64%—

Knowledge GPT-6 Luna leads

GPT-6 Luna: 57.0 (#41), MiMo-V2-Omni: 40.5 (#118)

Knowledge benchmarks
BenchmarkGPT-6 LunaMiMo-V2-Omni
LMArena Expert14441449
GPQA Diamond90.5%—
SimpleQA Verified41.4%—

Multimodal GPT-6 Luna leads

GPT-6 Luna: 42.4 (#30), MiMo-V2-Omni: 38.6 (#63)

Multimodal benchmarks
BenchmarkGPT-6 LunaMiMo-V2-Omni
LMArena Vision12171228
Blueprint-Bench 231.2%—
Furniture Assembly44.2%—

Multilingual MiMo-V2-Omni leads

GPT-6 Luna: 50.5 (#117), MiMo-V2-Omni: 51.8 (#102)

Multilingual benchmarks
BenchmarkGPT-6 LunaMiMo-V2-Omni
LMArena Non-English13861404
LMArena Chinese14331465
LMArena French14201447
LMArena German13691399
LMArena Japanese13691317
LMArena Korean13601355
LMArena Russian13941412
LMArena Spanish13931434

Instruction Following Too close to call

GPT-6 Luna: 74.3 (#99), MiMo-V2-Omni: 75.2 (#66)

Instruction Following benchmarks
BenchmarkGPT-6 LunaMiMo-V2-Omni
LMArena Instruction Following14091428

Long Context MiMo-V2-Omni leads

GPT-6 Luna: 43.0 (#111), MiMo-V2-Omni: 44.1 (#76)

Long Context benchmarks
BenchmarkGPT-6 LunaMiMo-V2-Omni
LMArena Longer Query14091442

Writing & Preference MiMo-V2-Omni leads

GPT-6 Luna: 58.3 (#119), MiMo-V2-Omni: 61.4 (#87)

Writing & Preference benchmarks
BenchmarkGPT-6 LunaMiMo-V2-Omni
LMArena Text13911423
LMArena Creative Writing13631392
LMArena Multi-Turn13961445

Frequently asked questions

Is GPT-6 Luna better than MiMo-V2-Omni?

GPT-6 Luna is the stronger model overall, scoring 53.3 to 43.6 on the Noometry Index.

Which is cheaper, GPT-6 Luna or MiMo-V2-Omni?

MiMo-V2-Omni is cheaper. It lists at $0.14 per million input tokens and $0.28 per million output tokens; GPT-6 Luna lists at $0.10 and $0.50.

Is GPT-6 Luna or MiMo-V2-Omni better for coding?

GPT-6 Luna scores higher on coding benchmarks: 55.5 versus 43.3 in the Noometry coding category.

Which has the bigger context window?

GPT-6 Luna does, with 1.05M tokens against 262K.

How many benchmarks do GPT-6 Luna and MiMo-V2-Omni share?

18 benchmarks have published results for both models. GPT-6 Luna has 42 scored results on Noometry and MiMo-V2-Omni has 18.

Related comparisons

Go deeper