Model comparison

Qwen-7B vs Qwen2.5-Coder (1.5B)

Neither Qwen-7B nor Qwen2.5-Coder (1.5B) has enough public benchmark results to be ranked yet; the rows below show what has been published.

Last verified . 4 shared benchmarks.

Summary

  • They share 4 benchmarks with published results for both.

Side by side

Qwen-7B and Qwen2.5-Coder (1.5B) specifications
Qwen-7BQwen2.5-Coder (1.5B)
ProviderAlibaba (Qwen)Alibaba (Qwen)
Noometry Index——
Released2023-09-282024-09-18
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked86

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Reasoning Not comparable

Qwen-7B: —, Qwen2.5-Coder (1.5B): —

Reasoning benchmarks
BenchmarkQwen-7BQwen2.5-Coder (1.5B)
Epoch Capabilities Index106.83113.14
BIG-Bench Hard45%—
HellaSwag—76.8%
LAMBADA67.9%—
PIQA77.9%—
WinoGrande—72.9%

Math Not comparable

Qwen-7B: —, Qwen2.5-Coder (1.5B): —

Math benchmarks
BenchmarkQwen-7BQwen2.5-Coder (1.5B)
GSM8K51.7%86.7%

Knowledge Not comparable

Qwen-7B: —, Qwen2.5-Coder (1.5B): —

Knowledge benchmarks
BenchmarkQwen-7BQwen2.5-Coder (1.5B)
ARC (AI2) Challenge75.3%60.9%
MMLU45%68%
BoolQ76.4%—

Frequently asked questions

Is Qwen-7B better than Qwen2.5-Coder (1.5B)?

Neither Qwen-7B nor Qwen2.5-Coder (1.5B) has enough public benchmark results to be ranked yet; the rows below show what has been published.

How many benchmarks do Qwen-7B and Qwen2.5-Coder (1.5B) share?

4 benchmarks have published results for both models. Qwen-7B has 8 scored results on Noometry and Qwen2.5-Coder (1.5B) has 6.

Related comparisons

Go deeper