Alibaba (Qwen), open weights

# Qwen-14B

> Qwen-14B by Alibaba (Qwen), released September 2023. Ranked #275 of 354 with a Noometry Index of 31.4. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/qwen-14b
- Last updated: 2026-10-10
- Title: Qwen-14B Benchmarks, Price & Rank (October 2026) | Noometry

Qwen-14B by Alibaba (Qwen) ranks 275th of 354 ranked models on the Noometry Index as of October 2026, with a score of 31.4. Its strongest category is math, where it ranks 227th.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #275 of 354
- **Index score:** 31.4
- **Evidence:** Confirmed 18 results
- **Provider:** [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba)
- **Released:** September 24, 2023
- **Weights:** Open weights
- **Reasoning:** Unknown
- **Context window:** —
- **Max output:** —
- **Input price:** Not listed
- **Output price:** Not listed
- **Blended price:** Not listed
- **Output speed:** Not measured
- **Value:** Not ranked
- **Knowledge cutoff:** Unknown

## Category scores

Each category score combines every public result we have in that category.

Qwen-14B category scores

1.  Coding 31.2
2.  Reasoning 19.6
3.  Math 31.2
4.  Multilingual 27.5
5.  Instruction Following 52.4
6.  Long Context 31.3
7.  Writing & Preference 27.6
8.  0204060

Qwen-14B category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 31.2 | #288 | 1 |
| [Reasoning](https://noometry.com/best/reasoning) | 19.6 | #257 | 1 |
| [Math](https://noometry.com/best/math) | 31.2 | #227 | 1 |
| [Multilingual](https://noometry.com/best/multilingual) | 27.5 | #275 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 52.4 | #289 | 1 |
| [Long Context](https://noometry.com/best/long-context) | 31.3 | #280 | 1 |
| [Writing & Preference](https://noometry.com/best/writing) | 27.6 | #299 | 3 |

## Strengths and weaknesses

Categories where Qwen-14B places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Qwen-14B: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Math](https://noometry.com/best/math) | 31.2 | −5.4 | #227 of 327, top 70% |
| [Reasoning](https://noometry.com/best/reasoning) | 19.6 | −4.1 | #257 of 350, top 74% |
| [Coding](https://noometry.com/best/coding) | 31.2 | −7.6 | #288 of 340, top 85% |

### Weakest categories

Qwen-14B: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Writing & Preference](https://noometry.com/best/writing) | 27.6 | −26.1 | #299 of 312, top 96% |
| [Instruction Following](https://noometry.com/best/instruction-following) | 52.4 | −18.9 | #289 of 305, top 95% |
| [Long Context](https://noometry.com/best/long-context) | 31.3 | −9.7 | #280 of 296, top 95% |

## Closest competitors

The models ranked just above and below Qwen-14B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen-14B
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [o1-pro](https://noometry.com/models/o1-pro) | #271 | 31.5 | $263 | — | [Compare](https://noometry.com/compare/o1-pro-vs-qwen-14b) |
| [Command R](https://noometry.com/models/command-r) | #272 | 31.4 | $0.26 | — | [Compare](https://noometry.com/compare/command-r-vs-qwen-14b) |
| [Qwen1.5-7B](https://noometry.com/models/qwen1-5-7b) | #273 | 31.4 | — | — | [Compare](https://noometry.com/compare/qwen-14b-vs-qwen1-5-7b) |
| [Wizardlm 13b](https://noometry.com/models/wizardlm-13b) | #274 | 31.4 | — | — | [Compare](https://noometry.com/compare/qwen-14b-vs-wizardlm-13b) |
| [Phi 3 Mini 4k Instruct June 2024](https://noometry.com/models/phi-3-mini-4k-instruct-june) | #276 | 31.3 | — | — | [Compare](https://noometry.com/compare/phi-3-mini-4k-instruct-june-vs-qwen-14b) |
| [Gemma 1.1 7b IT](https://noometry.com/models/gemma-1-1-7b-it) | #277 | 31.3 | — | — | [Compare](https://noometry.com/compare/gemma-1-1-7b-it-vs-qwen-14b) |
| [Mistral Small 3](https://noometry.com/models/mistral-small-3) | #278 | 31.2 | $0.0575 | — | [Compare](https://noometry.com/compare/mistral-small-3-vs-qwen-14b) |
| [Phi-4](https://noometry.com/models/phi-4) | #279 | 31.2 | $0.0875 | — | [Compare](https://noometry.com/compare/phi-4-vs-qwen-14b) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Qwen-14B Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1071 | #279 of 294, top 95% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Reasoning

Qwen-14B Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1027 | #287 of 297, top 97% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [BIG-Bench Hard](https://noometry.com/benchmarks/bbh) | 55% | #18 of 27, top 67% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [BIG-Bench Hard](https://noometry.com/benchmarks/bbh) | 53.4% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Epoch Capabilities Index](https://noometry.com/benchmarks/epoch-capabilities-index) | 113.03 | #187 of 213, top 88% |  | [Epoch AI](https://epoch.ai/eci) | 2023-09-24 |
| [LAMBADA](https://noometry.com/benchmarks/lambada) | 71.1% | #8 of 9, top 89% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [PIQA](https://noometry.com/benchmarks/piqa) | 79.9% | #23 of 27, top 86% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Math

Qwen-14B Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1068 | #273 of 285, top 96% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [GSM8K](https://noometry.com/benchmarks/gsm8k) | 61.2% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [GSM8K](https://noometry.com/benchmarks/gsm8k) | 61.3% | #16 of 38, top 43% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Knowledge

Qwen-14B Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [ARC (AI2) Challenge](https://noometry.com/benchmarks/arc-challenge) | 84.4% | #11 of 39, top 29% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [BoolQ](https://noometry.com/benchmarks/boolq) | 86.2% | #6 of 23, top 27% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [MMLU](https://noometry.com/benchmarks/mmlu) | 66.3% | #53 of 81, top 66% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [MMLU](https://noometry.com/benchmarks/mmlu) | 65% |  |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Multilingual

Qwen-14B Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1041 | #275 of 297, top 93% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1077 | #258 of 285, top 91% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Qwen-14B Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1031 | #285 of 298, top 96% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

Qwen-14B Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1028 | #282 of 291, top 97% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Qwen-14B Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1051 | #289 of 297, top 98% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1028 | #286 of 295, top 97% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1022 | #283 of 295, top 96% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

## Compare Qwen-14B

-   [Qwen-14B vs Wizardlm 13b](https://noometry.com/compare/qwen-14b-vs-wizardlm-13b)
-   [Qwen-14B vs Phi 3 Mini 4k Instruct June 2024](https://noometry.com/compare/phi-3-mini-4k-instruct-june-vs-qwen-14b)
-   [Qwen-14B vs Qwen1.5-7B](https://noometry.com/compare/qwen-14b-vs-qwen1-5-7b)
-   [Qwen-14B vs Gemma 1.1 7b IT](https://noometry.com/compare/gemma-1-1-7b-it-vs-qwen-14b)
-   [Qwen-14B vs Command R](https://noometry.com/compare/command-r-vs-qwen-14b)
-   [Qwen-14B vs Mistral Small 3](https://noometry.com/compare/mistral-small-3-vs-qwen-14b)
-   [Qwen-14B vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-qwen-14b)
-   [Qwen-14B vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-qwen-14b)
-   [Qwen-14B vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-qwen-14b)
-   [Qwen-14B vs Kimi K3](https://noometry.com/compare/kimi-k3-vs-qwen-14b)
-   [Qwen-14B vs Grok 4.6](https://noometry.com/compare/grok-4-6-vs-qwen-14b)
-   [Qwen-14B vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-qwen-14b)
-   [Qwen-14B vs Muse Spark 1.3](https://noometry.com/compare/muse-spark-1-3-vs-qwen-14b)
-   [Qwen-14B vs DeepSeek V4 Pro](https://noometry.com/compare/deepseek-v4-pro-vs-qwen-14b)

## Other Alibaba (Qwen) models

-   [Qwen3.8 Max](https://noometry.com/models/qwen3-8-max)56.8
-   [Qwen3.7 Max](https://noometry.com/models/qwen3-7-max)51.5
-   [Qwen3.6 Max Preview](https://noometry.com/models/qwen3-6-max-preview)51.5
-   [Qwen3.6 Plus](https://noometry.com/models/qwen3-6-plus)47.5
-   [Qwen3.5 397B-A17B](https://noometry.com/models/qwen3-5-397b-a17b)46.0
-   [Qwen3.8 27B](https://noometry.com/models/qwen3-8-27b)46.0
-   [Qwen3.5 Max Preview](https://noometry.com/models/qwen3-5-max-preview)45.3
-   [Qwen3.7 Plus](https://noometry.com/models/qwen3-7-plus)45.3

## Frequently asked questions

### How good is Qwen-14B?

Qwen-14B by Alibaba (Qwen) ranks 275th of 354 ranked models on the Noometry Index as of October 2026, with a score of 31.4. Its strongest category is math, where it ranks 227th.

### Is Qwen-14B open source?

Yes. Qwen-14B's weights are downloadable; check the license for commercial terms.

### What are Qwen-14B's strengths and weaknesses?

Relative to other ranked models, Qwen-14B places best in math, reasoning, coding and lowest in writing & preference, instruction following, long context.

### What is Qwen-14B best at?

Its best category is math, where it ranks 227th on Noometry.

### Cite this page

Noometry. (2026). Qwen-14B benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/qwen-14b

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/qwen-14b.md).
