Alibaba (Qwen), open weights

Qwen-14B

Qwen-14B by Alibaba (Qwen) ranks 275th of 354 ranked models on the Noometry Index as of October 2026, with a score of 31.4. Its strongest category is math, where it ranks 227th.

Last verified

Specifications

Noometry rank
#275 of 354
Index score
31.4
Evidence
Confirmed 18 results
Released
September 24, 2023
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Qwen-14B category scores
  1. Coding 31.2
  2. Reasoning 19.6
  3. Math 31.2
  4. Multilingual 27.5
  5. Instruction Following 52.4
  6. Long Context 31.3
  7. Writing & Preference 27.6
Qwen-14B category ranks
CategoryScoreRankResults
Coding31.2#2881
Reasoning19.6#2571
Math31.2#2271
Multilingual27.5#2751
Instruction Following52.4#2891
Long Context31.3#2801
Writing & Preference27.6#2993

Strengths and weaknesses

Categories where Qwen-14B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Qwen-14B: strongest categories
CategoryScorevs medianRank
Math31.2−5.4#227 of 327, top 70%
Reasoning19.6−4.1#257 of 350, top 74%
Coding31.2−7.6#288 of 340, top 85%

Weakest categories

Qwen-14B: weakest categories
CategoryScorevs medianRank
Writing & Preference27.6−26.1#299 of 312, top 96%
Instruction Following52.4−18.9#289 of 305, top 95%
Long Context31.3−9.7#280 of 296, top 95%

Closest competitors

The models ranked just above and below Qwen-14B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen-14B
ModelRankScoreBlended $/MSpeed
o1-pro#27131.5$263—Compare
Command R#27231.4$0.26—Compare
Qwen1.5-7B#27331.4——Compare
Wizardlm 13b#27431.4——Compare
Phi 3 Mini 4k Instruct June 2024#27631.3——Compare
Gemma 1.1 7b IT#27731.3——Compare
Mistral Small 3#27831.2$0.0575—Compare
Phi-4#27931.2$0.0875—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Qwen-14B Coding benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Coding1071#279 of 294, top 95%LMArena2026-10-08

Reasoning

Qwen-14B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Hard Prompts1027#287 of 297, top 97%LMArena2026-10-08
BIG-Bench Hard55%#18 of 27, top 67%Epoch AI
BIG-Bench Hard53.4%Epoch AI
Epoch Capabilities Index113.03#187 of 213, top 88%Epoch AI2023-09-24
LAMBADA71.1%#8 of 9, top 89%Epoch AI
PIQA79.9%#23 of 27, top 86%Epoch AI

Math

Qwen-14B Math benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Math1068#273 of 285, top 96%LMArena2026-10-08
GSM8K61.2%Epoch AI
GSM8K61.3%#16 of 38, top 43%Epoch AI

Knowledge

Qwen-14B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
ARC (AI2) Challenge84.4%#11 of 39, top 29%Epoch AI
BoolQ86.2%#6 of 23, top 27%Epoch AI
MMLU66.3%#53 of 81, top 66%Epoch AI
MMLU65%Epoch AI

Multilingual

Qwen-14B Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1041#275 of 297, top 93%LMArena2026-10-08
LMArena Chinese1077#258 of 285, top 91%LMArena2026-10-08

Instruction Following

Qwen-14B Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1031#285 of 298, top 96%LMArena2026-10-08

Long Context

Qwen-14B Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1028#282 of 291, top 97%LMArena2026-10-08

Writing & Preference

Qwen-14B Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1051#289 of 297, top 98%LMArena2026-10-08
LMArena Creative Writing1028#286 of 295, top 97%LMArena2026-10-08
LMArena Multi-Turn1022#283 of 295, top 96%LMArena2026-10-08

Compare Qwen-14B

Other Alibaba (Qwen) models

Frequently asked questions

How good is Qwen-14B?

Qwen-14B by Alibaba (Qwen) ranks 275th of 354 ranked models on the Noometry Index as of October 2026, with a score of 31.4. Its strongest category is math, where it ranks 227th.

Is Qwen-14B open source?

Yes. Qwen-14B's weights are downloadable; check the license for commercial terms.

What are Qwen-14B's strengths and weaknesses?

Relative to other ranked models, Qwen-14B places best in math, reasoning, coding and lowest in writing & preference, instruction following, long context.

What is Qwen-14B best at?

Its best category is math, where it ranks 227th on Noometry.