Model comparison
Mistral Small 3.2 vs Qwen2.5-VL 72B Instruct
Mistral Small 3.2 is the stronger model overall, scoring 31.2 to 29.9 on the Noometry Index.
Last verified . 1 shared benchmarks.
Summary
- They share 1 benchmark with published results for both. Mistral Small 3.2 scores higher in 0 categories and Qwen2.5-VL 72B Instruct in 1 category; one gap is clear of the uncertainty.
- Mistral Small 3.2 is cheaper at $0.0938 / $0.25 per million input/output tokens, against $2.80 / $8.40 for Qwen2.5-VL 72B Instruct.
- Mistral Small 3.2 accepts more context: 256K tokens versus 131K.
Side by side
| Mistral Small 3.2 | Qwen2.5-VL 72B Instruct | |
|---|---|---|
| Provider | Mistral AI | Alibaba (Qwen) |
| Noometry Index | 31.2 | 29.9 |
| Released | 2025-06-20 | 2024-09 |
| Weights | Open | Open |
| Context window | 256K | 131K |
| Max output | 16K | 8K |
| Input $ / M tokens | $0.0938 | $2.80 |
| Output $ / M tokens | $0.25 | $8.40 |
| Results tracked | 6 | 6 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Agentic & Tool Use Not comparable
Mistral Small 3.2: —, Qwen2.5-VL 72B Instruct: 18.6 (#144)
| Benchmark | Mistral Small 3.2 | Qwen2.5-VL 72B Instruct |
|---|---|---|
| OSWorld | — | 5% |
Reasoning Qwen2.5-VL 72B Instruct leads
Mistral Small 3.2: 18.1 (#287), Qwen2.5-VL 72B Instruct: 20.7 (#233)
| Benchmark | Mistral Small 3.2 | Qwen2.5-VL 72B Instruct |
|---|---|---|
| Kagi LLM Benchmark | 40.4% | 36% |
| Chess Puzzles | 1% | — |
| Epoch Capabilities Index | 131.74 | — |
Math Not comparable
Mistral Small 3.2: 26.3 (#260), Qwen2.5-VL 72B Instruct: —
| Benchmark | Mistral Small 3.2 | Qwen2.5-VL 72B Instruct |
|---|---|---|
| OTIS Mock AIME 2024-2025 | 30.3% | — |
Knowledge Not comparable
Mistral Small 3.2: 26.7 (#256), Qwen2.5-VL 72B Instruct: —
| Benchmark | Mistral Small 3.2 | Qwen2.5-VL 72B Instruct |
|---|---|---|
| GPQA Diamond | 49.1% | — |
Multimodal Not comparable
Mistral Small 3.2: —, Qwen2.5-VL 72B Instruct: 33.5 (#97)
| Benchmark | Mistral Small 3.2 | Qwen2.5-VL 72B Instruct |
|---|---|---|
| LMArena Vision | — | 1107 |
| Video-MME | — | 73.5% |
| GeoBench | — | 62% |
| SpatialViz-Bench | — | 33.3% |
Writing & Preference Not comparable
Mistral Small 3.2: 45.0 (#224), Qwen2.5-VL 72B Instruct: —
| Benchmark | Mistral Small 3.2 | Qwen2.5-VL 72B Instruct |
|---|---|---|
| EQ-Bench Creative Writing | 1255 | — |
Frequently asked questions
Is Mistral Small 3.2 better than Qwen2.5-VL 72B Instruct?
Mistral Small 3.2 is the stronger model overall, scoring 31.2 to 29.9 on the Noometry Index.
Which is cheaper, Mistral Small 3.2 or Qwen2.5-VL 72B Instruct?
Mistral Small 3.2 is cheaper. It lists at $0.0938 per million input tokens and $0.25 per million output tokens; Qwen2.5-VL 72B Instruct lists at $2.80 and $8.40.
Which has the bigger context window?
Mistral Small 3.2 does, with 256K tokens against 131K.
How many benchmarks do Mistral Small 3.2 and Qwen2.5-VL 72B Instruct share?
1 benchmark has published results for both models. Mistral Small 3.2 has 6 scored results on Noometry and Qwen2.5-VL 72B Instruct has 6.