Model comparison

Mixtral 8x7B vs Qwen3-1.7B

Mixtral 8x7B and Qwen3-1.7B score almost the same on the Noometry Index (27.1 vs 26.6), so choose on price, context window or the category you care about most.

Last verified . 1 shared benchmarks.

Mixtral 8x7B Mistral AI

27.1

Rank #334 Confirmed

Qwen3-1.7B Alibaba (Qwen)

26.6

Rank #336 Reported

Summary

  • They share 1 benchmark with published results for both. Mixtral 8x7B scores higher in 1 category and Qwen3-1.7B in 2 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Qwen3-1.7B leads 19.6 to 11.0.
  • The biggest single-benchmark swing is GPQA Diamond: 30.6% for Mixtral 8x7B and 38% for Qwen3-1.7B.

Side by side

Mixtral 8x7B and Qwen3-1.7B specifications
Mixtral 8x7BQwen3-1.7B
ProviderMistral AIAlibaba (Qwen)
Noometry Index27.126.6
Released2023-12-112025-04-29
WeightsOpenOpen
Context window32K—
Max output32K—
Input $ / M tokens$0.70—
Output $ / M tokens$0.70—
Results tracked384

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mixtral 8x7B: 32.8 (#269), Qwen3-1.7B: —

Coding benchmarks
BenchmarkMixtral 8x7BQwen3-1.7B
LMArena Coding1126—
HumanEval+39.6%—
MBPP+49.7%—

Agentic & Tool Use Not comparable

Mixtral 8x7B: —, Qwen3-1.7B: 24.7 (#115)

Agentic & Tool Use benchmarks
BenchmarkMixtral 8x7BQwen3-1.7B
Berkeley Function Calling Leaderboard—28.4%

Reasoning Qwen3-1.7B leads

Mixtral 8x7B: 18.2 (#285), Qwen3-1.7B: 19.2 (#267)

Reasoning benchmarks
BenchmarkMixtral 8x7BQwen3-1.7B
Chess Puzzles—0%
LMArena Hard Prompts1115—
DTBench49.6%—
Adversarial NLI55.2%—
Epoch Capabilities Index118.47—
ForecastBench56.3—
HellaSwag86.7%—
PIQA83.6%—
WinoGrande77.2%—

Math Mixtral 8x7B leads

Mixtral 8x7B: 18.8 (#289), Qwen3-1.7B: 16.3 (#294)

Math benchmarks
BenchmarkMixtral 8x7BQwen3-1.7B
OTIS Mock AIME 2024-2025—8.1%
Omni-MATH10.5%—
LMArena Math1147—
MATH Level 510%—
GSM8K74.4%—

Knowledge Qwen3-1.7B leads

Mixtral 8x7B: 11.0 (#301), Qwen3-1.7B: 19.6 (#278)

Knowledge benchmarks
BenchmarkMixtral 8x7BQwen3-1.7B
GPQA Diamond30.6%38%
MMLU-Pro33.5%—
GPQA (HELM)29.6%—
LMArena Expert1088—
ARC (AI2) Challenge87.3%—
MMLU70.6%—
OpenBookQA85.8%—
TriviaQA82.2%—

Multilingual Not comparable

Mixtral 8x7B: 29.6 (#266), Qwen3-1.7B: —

Multilingual benchmarks
BenchmarkMixtral 8x7BQwen3-1.7B
LMArena Non-English1077—
LMArena Chinese1055—
LMArena French1166—
LMArena German1114—
LMArena Japanese931—
LMArena Korean968—
LMArena Russian1090—
LMArena Spanish1111—

Instruction Following Not comparable

Mixtral 8x7B: 51.0 (#297), Qwen3-1.7B: —

Instruction Following benchmarks
BenchmarkMixtral 8x7BQwen3-1.7B
IFEval57.5%—
LMArena Instruction Following1109—

Long Context Not comparable

Mixtral 8x7B: 33.4 (#260), Qwen3-1.7B: —

Long Context benchmarks
BenchmarkMixtral 8x7BQwen3-1.7B
LMArena Longer Query1103—

Writing & Preference Not comparable

Mixtral 8x7B: 34.2 (#270), Qwen3-1.7B: —

Writing & Preference benchmarks
BenchmarkMixtral 8x7BQwen3-1.7B
LMArena Text1132—
LMArena Creative Writing1109—
WildBench67.3%—
LMArena Multi-Turn1115—

Frequently asked questions

Is Mixtral 8x7B better than Qwen3-1.7B?

Mixtral 8x7B and Qwen3-1.7B score almost the same on the Noometry Index (27.1 vs 26.6), so choose on price, context window or the category you care about most.

How many benchmarks do Mixtral 8x7B and Qwen3-1.7B share?

1 benchmark has published results for both models. Mixtral 8x7B has 38 scored results on Noometry and Qwen3-1.7B has 4.

Related comparisons

Go deeper