Model comparison

Laguna M.1 vs Qwen Plus

Qwen Plus is the stronger model overall, scoring 37.1 to 32.5 on the Noometry Index.

Last verified . 0 shared benchmarks.

Laguna M.1 Poolside

32.5

Rank #256 Reported

Qwen Plus Alibaba (Qwen)

37.1

Rank #210 Confirmed

Summary

  • The widest gap is in reasoning, where Qwen Plus leads 28.4 to 23.1.
  • Qwen Plus accepts more context: 1M tokens versus 262K.
  • Laguna M.1 has downloadable open weights; the other is API-only.

Side by side

Laguna M.1 and Qwen Plus specifications
Laguna M.1Qwen Plus
ProviderPoolsideAlibaba (Qwen)
Noometry Index32.537.1
Released2026-04-282024-01-25
WeightsOpenProprietary
Context window262K1M
Max output33K33K
Input $ / M tokens—$0.40
Output $ / M tokens—$1.20
Results tracked320

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen Plus leads

Laguna M.1: 36.6 (#204), Qwen Plus: 38.9 (#167)

Coding benchmarks
BenchmarkLaguna M.1Qwen Plus
LMArena WebDev1349—
LMArena Coding—1328

Reasoning Qwen Plus leads

Laguna M.1: 23.1 (#184), Qwen Plus: 28.4 (#107)

Reasoning benchmarks
BenchmarkLaguna M.1Qwen Plus
Kagi LLM Benchmark—63.3%
LMArena Hard Prompts—1317
DTBench—81.1%
LMCA—24%
Surface Evolver Bench15.6%—

Math Qwen Plus leads

Laguna M.1: 21.1 (#283), Qwen Plus: 23.3 (#271)

Math benchmarks
BenchmarkLaguna M.1Qwen Plus
OTIS Mock AIME 2024-2025—17.8%
ProofBench0%—
LMArena Math—1326
MATH Level 5—65.3%
FrontierMath (Feb 2025 set)—1.7%

Knowledge Not comparable

Laguna M.1: —, Qwen Plus: 27.4 (#251)

Knowledge benchmarks
BenchmarkLaguna M.1Qwen Plus
GPQA Diamond—48.1%
LMArena Expert—1328

Multilingual Not comparable

Laguna M.1: —, Qwen Plus: 45.1 (#175)

Multilingual benchmarks
BenchmarkLaguna M.1Qwen Plus
LMArena Non-English—1310
LMArena Chinese—1347
LMArena Japanese—1251
LMArena Russian—1323

Instruction Following Not comparable

Laguna M.1: —, Qwen Plus: 68.8 (#181)

Instruction Following benchmarks
BenchmarkLaguna M.1Qwen Plus
LMArena Instruction Following—1303

Long Context Not comparable

Laguna M.1: —, Qwen Plus: 40.3 (#158)

Long Context benchmarks
BenchmarkLaguna M.1Qwen Plus
LMArena Longer Query—1324

Writing & Preference Not comparable

Laguna M.1: —, Qwen Plus: 52.2 (#176)

Writing & Preference benchmarks
BenchmarkLaguna M.1Qwen Plus
LMArena Text—1326
LMArena Creative Writing—1293
LMArena Multi-Turn—1336

Frequently asked questions

Is Laguna M.1 better than Qwen Plus?

Qwen Plus is the stronger model overall, scoring 37.1 to 32.5 on the Noometry Index.

Is Laguna M.1 or Qwen Plus better for coding?

Qwen Plus scores higher on coding benchmarks: 38.9 versus 36.6 in the Noometry coding category.

Which has the bigger context window?

Qwen Plus does, with 1M tokens against 262K.

How many benchmarks do Laguna M.1 and Qwen Plus share?

0 benchmarks have published results for both models. Laguna M.1 has 3 scored results on Noometry and Qwen Plus has 20.

Related comparisons

Go deeper