Model comparison

Phi-1.5 vs Qwen2.5-Coder (1.5B)

Neither Phi-1.5 nor Qwen2.5-Coder (1.5B) has enough public benchmark results to be ranked yet; the rows below show what has been published.

Last verified . 5 shared benchmarks.

Summary

  • They share 5 benchmarks with published results for both.

Side by side

Phi-1.5 and Qwen2.5-Coder (1.5B) specifications
Phi-1.5Qwen2.5-Coder (1.5B)
ProviderMicrosoftAlibaba (Qwen)
Noometry Index——
Released2023-09-112024-09-18
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked76

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Reasoning Not comparable

Phi-1.5: —, Qwen2.5-Coder (1.5B): —

Reasoning benchmarks
BenchmarkPhi-1.5Qwen2.5-Coder (1.5B)
Epoch Capabilities Index91.53113.14
HellaSwag47.6%76.8%
WinoGrande73.4%72.9%

Math Not comparable

Phi-1.5: —, Qwen2.5-Coder (1.5B): —

Math benchmarks
BenchmarkPhi-1.5Qwen2.5-Coder (1.5B)
GSM8K—86.7%

Knowledge Not comparable

Phi-1.5: —, Qwen2.5-Coder (1.5B): —

Knowledge benchmarks
BenchmarkPhi-1.5Qwen2.5-Coder (1.5B)
ARC (AI2) Challenge44.4%60.9%
MMLU37.6%68%
BoolQ75.8%—
OpenBookQA37.2%—

Frequently asked questions

Is Phi-1.5 better than Qwen2.5-Coder (1.5B)?

Neither Phi-1.5 nor Qwen2.5-Coder (1.5B) has enough public benchmark results to be ranked yet; the rows below show what has been published.

How many benchmarks do Phi-1.5 and Qwen2.5-Coder (1.5B) share?

5 benchmarks have published results for both models. Phi-1.5 has 7 scored results on Noometry and Qwen2.5-Coder (1.5B) has 6.

Related comparisons

Go deeper