Model comparison

GLM-5V-Turbo vs Mistral Nemo

GLM-5V-Turbo is the stronger model overall, scoring 43.8 to 26.4 on the Noometry Index. Mistral Nemo costs 13× less per token, which makes it the better buy when GLM-5V-Turbo's lead doesn't matter for your workload.

Last verified . 0 shared benchmarks.

GLM-5V-Turbo Z.ai (Zhipu)

43.8

Rank #84 Confirmed

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Summary

  • The widest gap is in writing & preference, where GLM-5V-Turbo leads 62.5 to 28.5.
  • Mistral Nemo is cheaper at $0.15 / $0.15 per million input/output tokens, against $1.20 / $4 for GLM-5V-Turbo.
  • GLM-5V-Turbo accepts more context: 200K tokens versus 128K.
  • Mistral Nemo has downloadable open weights; the other is API-only.

Side by side

GLM-5V-Turbo and Mistral Nemo specifications
GLM-5V-TurboMistral Nemo
ProviderZ.ai (Zhipu)Mistral AI
Noometry Index43.826.4
Released2026-04-012024-07-01
WeightsProprietaryOpen
Context window200K128K
Max output131K128K
Input $ / M tokens$1.20$0.15
Output $ / M tokens$4$0.15
Results tracked1910

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

GLM-5V-Turbo: 42.1 (#111), Mistral Nemo: —

Coding benchmarks
BenchmarkGLM-5V-TurboMistral Nemo
LMArena WebDev1401—
LMArena Coding1466—

Agentic & Tool Use Not comparable

GLM-5V-Turbo: —, Mistral Nemo: 23.5 (#125)

Agentic & Tool Use benchmarks
BenchmarkGLM-5V-TurboMistral Nemo
Berkeley Function Calling Leaderboard—27.6%
BALROG—17.6%

Reasoning GLM-5V-Turbo leads

GLM-5V-Turbo: 29.7 (#89), Mistral Nemo: 20.7 (#232)

Reasoning benchmarks
BenchmarkGLM-5V-TurboMistral Nemo
LMArena Hard Prompts1443—
DTBench—48.6%
Epoch Capabilities Index—118.68
PIQA—83.5%

Math GLM-5V-Turbo leads

GLM-5V-Turbo: 39.4 (#106), Mistral Nemo: 25.5 (#268)

Math benchmarks
BenchmarkGLM-5V-TurboMistral Nemo
LMArena Math1441—
MATH Level 5—10.8%
GSM8K—84.2%

Knowledge GLM-5V-Turbo leads

GLM-5V-Turbo: 40.6 (#117), Mistral Nemo: 12.3 (#298)

Knowledge benchmarks
BenchmarkGLM-5V-TurboMistral Nemo
GPQA Diamond—29.9%
LMArena Expert1452—
BoolQ—82.5%

Multimodal Not comparable

GLM-5V-Turbo: 40.9 (#42), Mistral Nemo: —

Multimodal benchmarks
BenchmarkGLM-5V-TurboMistral Nemo
LMArena Vision1264—
LMArena Document1416—

Multilingual Not comparable

GLM-5V-Turbo: 53.0 (#73), Mistral Nemo: —

Multilingual benchmarks
BenchmarkGLM-5V-TurboMistral Nemo
LMArena Non-English1420—
LMArena Chinese1488—
LMArena French1444—
LMArena German1423—
LMArena Korean1396—
LMArena Russian1431—
LMArena Spanish1450—

Instruction Following Not comparable

GLM-5V-Turbo: 75.0 (#80), Mistral Nemo: —

Instruction Following benchmarks
BenchmarkGLM-5V-TurboMistral Nemo
LMArena Instruction Following1423—

Long Context Not comparable

GLM-5V-Turbo: 44.0 (#80), Mistral Nemo: —

Long Context benchmarks
BenchmarkGLM-5V-TurboMistral Nemo
LMArena Longer Query1438—

Writing & Preference GLM-5V-Turbo leads

GLM-5V-Turbo: 62.5 (#73), Mistral Nemo: 28.5 (#296)

Writing & Preference benchmarks
BenchmarkGLM-5V-TurboMistral Nemo
LMArena Text1437—
LMArena Creative Writing1416—
EQ-Bench Creative Writing—881
LMArena Multi-Turn1432—

Frequently asked questions

Is GLM-5V-Turbo better than Mistral Nemo?

GLM-5V-Turbo is the stronger model overall, scoring 43.8 to 26.4 on the Noometry Index. Mistral Nemo costs 13× less per token, which makes it the better buy when GLM-5V-Turbo's lead doesn't matter for your workload.

Which is cheaper, GLM-5V-Turbo or Mistral Nemo?

Mistral Nemo is cheaper. It lists at $0.15 per million input tokens and $0.15 per million output tokens; GLM-5V-Turbo lists at $1.20 and $4.

Which has the bigger context window?

GLM-5V-Turbo does, with 200K tokens against 128K.

How many benchmarks do GLM-5V-Turbo and Mistral Nemo share?

0 benchmarks have published results for both models. GLM-5V-Turbo has 19 scored results on Noometry and Mistral Nemo has 10.

Related comparisons

Go deeper