Model comparison

Mistral Nemo vs Qwen Plus

Qwen Plus is the stronger model overall, scoring 37.1 to 26.4 on the Noometry Index. Mistral Nemo costs 4.0× less per token, which makes it the better buy when Qwen Plus's lead doesn't matter for your workload.

Last verified . 3 shared benchmarks.

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Qwen Plus Alibaba (Qwen)

37.1

Rank #210 Confirmed

Summary

  • They share 3 benchmarks with published results for both. Mistral Nemo scores higher in 1 category and Qwen Plus in 3 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen Plus leads 52.2 to 28.5.
  • The biggest single-benchmark swing is MATH Level 5: 10.8% for Mistral Nemo and 65.3% for Qwen Plus.
  • Mistral Nemo is cheaper at $0.15 / $0.15 per million input/output tokens, against $0.40 / $1.20 for Qwen Plus.
  • Qwen Plus accepts more context: 1M tokens versus 128K.
  • Mistral Nemo has downloadable open weights; the other is API-only.

Side by side

Mistral Nemo and Qwen Plus specifications
Mistral NemoQwen Plus
ProviderMistral AIAlibaba (Qwen)
Noometry Index26.437.1
Released2024-07-012024-01-25
WeightsOpenProprietary
Context window128K1M
Max output128K33K
Input $ / M tokens$0.15$0.40
Output $ / M tokens$0.15$1.20
Results tracked1020

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mistral Nemo: —, Qwen Plus: 38.9 (#167)

Coding benchmarks
BenchmarkMistral NemoQwen Plus
LMArena Coding—1328

Agentic & Tool Use Not comparable

Mistral Nemo: 23.5 (#125), Qwen Plus: —

Agentic & Tool Use benchmarks
BenchmarkMistral NemoQwen Plus
Berkeley Function Calling Leaderboard27.6%—
BALROG17.6%—

Reasoning Qwen Plus leads

Mistral Nemo: 20.7 (#232), Qwen Plus: 28.4 (#107)

Reasoning benchmarks
BenchmarkMistral NemoQwen Plus
DTBench48.6%81.1%
Kagi LLM Benchmark—63.3%
LMArena Hard Prompts—1317
LMCA—24%
Epoch Capabilities Index118.68—
PIQA83.5%—

Math Mistral Nemo leads

Mistral Nemo: 25.5 (#268), Qwen Plus: 23.3 (#271)

Math benchmarks
BenchmarkMistral NemoQwen Plus
MATH Level 510.8%65.3%
OTIS Mock AIME 2024-2025—17.8%
LMArena Math—1326
FrontierMath (Feb 2025 set)—1.7%
GSM8K84.2%—

Knowledge Qwen Plus leads

Mistral Nemo: 12.3 (#298), Qwen Plus: 27.4 (#251)

Knowledge benchmarks
BenchmarkMistral NemoQwen Plus
GPQA Diamond29.9%48.1%
LMArena Expert—1328
BoolQ82.5%—

Multilingual Not comparable

Mistral Nemo: —, Qwen Plus: 45.1 (#175)

Multilingual benchmarks
BenchmarkMistral NemoQwen Plus
LMArena Non-English—1310
LMArena Chinese—1347
LMArena Japanese—1251
LMArena Russian—1323

Instruction Following Not comparable

Mistral Nemo: —, Qwen Plus: 68.8 (#181)

Instruction Following benchmarks
BenchmarkMistral NemoQwen Plus
LMArena Instruction Following—1303

Long Context Not comparable

Mistral Nemo: —, Qwen Plus: 40.3 (#158)

Long Context benchmarks
BenchmarkMistral NemoQwen Plus
LMArena Longer Query—1324

Writing & Preference Qwen Plus leads

Mistral Nemo: 28.5 (#296), Qwen Plus: 52.2 (#176)

Writing & Preference benchmarks
BenchmarkMistral NemoQwen Plus
LMArena Text—1326
LMArena Creative Writing—1293
EQ-Bench Creative Writing881—
LMArena Multi-Turn—1336

Frequently asked questions

Is Mistral Nemo better than Qwen Plus?

Qwen Plus is the stronger model overall, scoring 37.1 to 26.4 on the Noometry Index. Mistral Nemo costs 4.0× less per token, which makes it the better buy when Qwen Plus's lead doesn't matter for your workload.

Which is cheaper, Mistral Nemo or Qwen Plus?

Mistral Nemo is cheaper. It lists at $0.15 per million input tokens and $0.15 per million output tokens; Qwen Plus lists at $0.40 and $1.20.

Which has the bigger context window?

Qwen Plus does, with 1M tokens against 128K.

How many benchmarks do Mistral Nemo and Qwen Plus share?

3 benchmarks have published results for both models. Mistral Nemo has 10 scored results on Noometry and Qwen Plus has 20.

Related comparisons

Go deeper