Model comparison

Mixtral 8x22B vs Qwen3-1.7B

Mixtral 8x22B and Qwen3-1.7B score almost the same on the Noometry Index (27.1 vs 26.6), so choose on price, context window or the category you care about most.

Last verified . 1 shared benchmarks.

Mixtral 8x22B Mistral AI

27.1

Rank #333 Confirmed

Qwen3-1.7B Alibaba (Qwen)

26.6

Rank #336 Reported

Summary

  • They share 1 benchmark with published results for both. Mixtral 8x22B scores higher in 2 categories and Qwen3-1.7B in 2 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in math, where Mixtral 8x22B leads 22.9 to 16.3.

Side by side

Mixtral 8x22B and Qwen3-1.7B specifications
Mixtral 8x22BQwen3-1.7B
ProviderMistral AIAlibaba (Qwen)
Noometry Index27.126.6
Released2024-04-172025-04-29
WeightsOpenOpen
Context window64K—
Max output64K—
Input $ / M tokens$2—
Output $ / M tokens$6—
Results tracked344

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mixtral 8x22B: 24.2 (#329), Qwen3-1.7B: —

Coding benchmarks
BenchmarkMixtral 8x22BQwen3-1.7B
WeirdML3.2%—
BigCodeBench Instruct40.6%—
LMArena Coding1166—
BigCodeBench Complete50.2%—
HumanEval+72%—
MBPP+64.3%—

Agentic & Tool Use Qwen3-1.7B leads

Mixtral 8x22B: 23.1 (#127), Qwen3-1.7B: 24.7 (#115)

Agentic & Tool Use benchmarks
BenchmarkMixtral 8x22BQwen3-1.7B
Berkeley Function Calling Leaderboard—28.4%
Cybench7.5%—

Reasoning Too close to call

Mixtral 8x22B: 19.9 (#248), Qwen3-1.7B: 19.2 (#267)

Reasoning benchmarks
BenchmarkMixtral 8x22BQwen3-1.7B
Chess Puzzles—0%
LMArena Hard Prompts1150—
DTBench55.1%—
Epoch Capabilities Index122.03—
ForecastBench56.3—

Math Mixtral 8x22B leads

Mixtral 8x22B: 22.9 (#275), Qwen3-1.7B: 16.3 (#294)

Math benchmarks
BenchmarkMixtral 8x22BQwen3-1.7B
OTIS Mock AIME 2024-2025—8.1%
Omni-MATH16.3%—
LMArena Math1184—
MATH Level 524.2%—

Knowledge Qwen3-1.7B leads

Mixtral 8x22B: 15.1 (#293), Qwen3-1.7B: 19.6 (#278)

Knowledge benchmarks
BenchmarkMixtral 8x22BQwen3-1.7B
GPQA Diamond34.1%38%
MMLU-Pro46%—
GPQA (HELM)33.4%—
LMArena Expert1113—
MMLU77.8%—

Multilingual Not comparable

Mixtral 8x22B: 32.8 (#255), Qwen3-1.7B: —

Multilingual benchmarks
BenchmarkMixtral 8x22BQwen3-1.7B
LMArena Non-English1128—
LMArena Chinese1116—
LMArena French1166—
LMArena German1141—
LMArena Japanese1037—
LMArena Korean1057—
LMArena Russian1158—
LMArena Spanish1151—

Instruction Following Not comparable

Mixtral 8x22B: 57.7 (#266), Qwen3-1.7B: —

Instruction Following benchmarks
BenchmarkMixtral 8x22BQwen3-1.7B
IFEval72.4%—
LMArena Instruction Following1147—

Long Context Not comparable

Mixtral 8x22B: 34.7 (#247), Qwen3-1.7B: —

Long Context benchmarks
BenchmarkMixtral 8x22BQwen3-1.7B
LMArena Longer Query1144—

Writing & Preference Not comparable

Mixtral 8x22B: 36.9 (#262), Qwen3-1.7B: —

Writing & Preference benchmarks
BenchmarkMixtral 8x22BQwen3-1.7B
LMArena Text1162—
LMArena Creative Writing1141—
WildBench71.1%—
LMArena Multi-Turn1130—

Frequently asked questions

Is Mixtral 8x22B better than Qwen3-1.7B?

Mixtral 8x22B and Qwen3-1.7B score almost the same on the Noometry Index (27.1 vs 26.6), so choose on price, context window or the category you care about most.

How many benchmarks do Mixtral 8x22B and Qwen3-1.7B share?

1 benchmark has published results for both models. Mixtral 8x22B has 34 scored results on Noometry and Qwen3-1.7B has 4.

Related comparisons

Go deeper