Alibaba (Qwen), open weights

Qwen1.5-14B

Qwen1.5-14B by Alibaba (Qwen) ranks 253rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.7. Its strongest category is math, where it ranks 215th.

Last verified

Specifications

Noometry rank
#253 of 354
Index score
32.7
Evidence
Confirmed 17 results
Released
February 4, 2024
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Qwen1.5-14B category scores
  1. Coding 33.1
  2. Reasoning 21.4
  3. Math 32.4
  4. Knowledge 29.8
  5. Multilingual 30.7
  6. Instruction Following 56.8
  7. Long Context 33.7
  8. Writing & Preference 33.6
Qwen1.5-14B category ranks
CategoryScoreRankResults
Coding33.1#2631
Reasoning21.4#2231
Math32.4#2151
Knowledge29.8#2321
Multilingual30.7#2621
Instruction Following56.8#2711
Long Context33.7#2571
Writing & Preference33.6#2763

Strengths and weaknesses

Categories where Qwen1.5-14B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Qwen1.5-14B: strongest categories
CategoryScorevs medianRank
Reasoning21.4−2.2#223 of 350, top 64%
Math32.4−4.2#215 of 327, top 66%
Knowledge29.8−7.5#232 of 314, top 74%

Weakest categories

Qwen1.5-14B: weakest categories
CategoryScorevs medianRank
Instruction Following56.8−14.4#271 of 305, top 89%
Writing & Preference33.6−20.2#276 of 312, top 89%
Multilingual30.7−16.7#262 of 297, top 89%

Closest competitors

The models ranked just above and below Qwen1.5-14B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen1.5-14B
ModelRankScoreBlended $/MSpeed
Wizardlm 70b#24933.0——Compare
Phi 3 Medium 4k Instruct#25033.0——Compare
Tulu 3 (Tülu 3) 70B#25133.0——Compare
DeepSeek-R1-Distill-Qwen-14B#25232.7——Compare
Olmo 2 0325 32b Instruct#25432.7——Compare
gpt-oss-20b#25532.5$0.03696Compare
Laguna M.1#25632.5——Compare
Command R+#25732.4$4.38—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Qwen1.5-14B Coding benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Coding1138#259 of 294, top 89%LMArena2026-10-08

Reasoning

Qwen1.5-14B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Hard Prompts1113#262 of 297, top 89%LMArena2026-10-08

Math

Qwen1.5-14B Math benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Math1125#260 of 285, top 92%LMArena2026-10-08

Knowledge

Qwen1.5-14B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Expert1094#251 of 273, top 92%LMArena2026-10-08
MMLU68.6%#51 of 81, top 63%Epoch AI

Multilingual

Qwen1.5-14B Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1095#262 of 297, top 89%LMArena2026-10-08
LMArena Chinese1147#247 of 285, top 87%LMArena2026-10-08
LMArena French1116#211 of 223, top 95%LMArena2026-10-08
LMArena German1043#221 of 231, top 96%LMArena2026-10-08
LMArena Japanese1019#197 of 211, top 94%LMArena2026-10-08
LMArena Russian1046#269 of 283, top 96%LMArena2026-10-08
LMArena Spanish1085#219 of 226, top 97%LMArena2026-10-08

Instruction Following

Qwen1.5-14B Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1102#267 of 298, top 90%LMArena2026-10-08

Long Context

Qwen1.5-14B Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1113#264 of 291, top 91%LMArena2026-10-08

Writing & Preference

Qwen1.5-14B Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1128#264 of 297, top 89%LMArena2026-10-08
LMArena Creative Writing1091#268 of 295, top 91%LMArena2026-10-08
LMArena Multi-Turn1110#263 of 295, top 90%LMArena2026-10-08

Compare Qwen1.5-14B

Other Alibaba (Qwen) models

Frequently asked questions

How good is Qwen1.5-14B?

Qwen1.5-14B by Alibaba (Qwen) ranks 253rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.7. Its strongest category is math, where it ranks 215th.

Is Qwen1.5-14B open source?

Yes. Qwen1.5-14B's weights are downloadable; check the license for commercial terms.

What are Qwen1.5-14B's strengths and weaknesses?

Relative to other ranked models, Qwen1.5-14B places best in reasoning, math, knowledge and lowest in instruction following, writing & preference, multilingual.

What is Qwen1.5-14B best at?

Its best category is math, where it ranks 215th on Noometry.