Meta, open weights

# Llama 13b

> Llama 13b by Meta, released February 2023. Ranked #348 of 354 with a Noometry Index of 24.4. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/llama-13b
- Last updated: 2026-10-10
- Title: Llama 13b Benchmarks, Price & Rank (October 2026) | Noometry

Llama 13b by Meta ranks 348th of 354 ranked models on the Noometry Index as of October 2026, with a score of 24.4. Its strongest category is math, where it ranks 256th.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #348 of 354
- **Index score:** 24.4
- **Evidence:** Confirmed 21 results
- **Provider:** [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta)
- **Released:** February 24, 2023
- **Weights:** Open weights
- **Reasoning:** Unknown
- **Context window:** —
- **Max output:** —
- **Input price:** Not listed
- **Output price:** Not listed
- **Blended price:** Not listed
- **Output speed:** Not measured
- **Value:** Not ranked
- **Knowledge cutoff:** Unknown

## Category scores

Each category score combines every public result we have in that category.

Llama 13b category scores

1.  Coding 21.4
2.  Reasoning 14.0
3.  Math 26.7
4.  Multilingual 16.6
5.  Instruction Following 36.7
6.  Writing & Preference 13.8
7.  010203040

Llama 13b category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 21.4 | #337 | 1 |
| [Reasoning](https://noometry.com/best/reasoning) | 14.0 | #329 | 1 |
| [Math](https://noometry.com/best/math) | 26.7 | #256 | 1 |
| [Multilingual](https://noometry.com/best/multilingual) | 16.6 | #297 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 36.7 | #305 | 1 |
| [Writing & Preference](https://noometry.com/best/writing) | 13.8 | #312 | 3 |

## Strengths and weaknesses

Categories where Llama 13b places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Llama 13b: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Math](https://noometry.com/best/math) | 26.7 | −9.9 | #256 of 327, top 79% |
| [Reasoning](https://noometry.com/best/reasoning) | 14.0 | −9.6 | #329 of 350, top 94% |
| [Coding](https://noometry.com/best/coding) | 21.4 | −17.4 | #337 of 340, top 100% |

### Weakest categories

Llama 13b: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Writing & Preference](https://noometry.com/best/writing) | 13.8 | −40.0 | #312 of 312, top 100% |
| [Instruction Following](https://noometry.com/best/instruction-following) | 36.7 | −34.6 | #305 of 305, top 100% |
| [Multilingual](https://noometry.com/best/multilingual) | 16.6 | −30.8 | #297 of 297, top 100% |

## Closest competitors

The models ranked just above and below Llama 13b. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Llama 13b
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [Llama 3-8B](https://noometry.com/models/llama-3-8b) | #344 | 25.5 | — | — | [Compare](https://noometry.com/compare/llama-13b-vs-llama-3-8b) |
| [Claude 2.1](https://noometry.com/models/claude-2-1) | #345 | 25.2 | — | — | [Compare](https://noometry.com/compare/claude-2-1-vs-llama-13b) |
| [Claude 2](https://noometry.com/models/claude-2) | #346 | 25.0 | — | — | [Compare](https://noometry.com/compare/claude-2-vs-llama-13b) |
| [DeepSeek LLM 67B](https://noometry.com/models/deepseek-llm-67b) | #347 | 24.9 | — | — | [Compare](https://noometry.com/compare/deepseek-llm-67b-vs-llama-13b) |
| [Llama 2-70B](https://noometry.com/models/llama-2-70b) | #349 | 24.4 | — | — | [Compare](https://noometry.com/compare/llama-13b-vs-llama-2-70b) |
| [GPT-3.5-turbo](https://noometry.com/models/gpt-3-5-turbo) | #350 | 23.2 | $0.75 | — | [Compare](https://noometry.com/compare/gpt-3-5-turbo-vs-llama-13b) |
| [Mistral 7B](https://noometry.com/models/mistral-7b) | #351 | 23.0 | $0.25 | — | [Compare](https://noometry.com/compare/llama-13b-vs-mistral-7b) |
| [Llama 3.1-8B](https://noometry.com/models/llama-3-1-8b) | #352 | 23.0 | $0.0575 | — | [Compare](https://noometry.com/compare/llama-13b-vs-llama-3-1-8b) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Llama 13b Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 683 | #294 of 294, top 100% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Reasoning

Llama 13b Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 728 | #297 of 297, top 100% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [BIG-Bench Hard](https://noometry.com/benchmarks/bbh) | 37.9% | #23 of 27, top 86% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Epoch Capabilities Index](https://noometry.com/benchmarks/epoch-capabilities-index) | 100.58 | #203 of 213, top 96% |  | [Epoch AI](https://epoch.ai/eci) | 2023-02-24 |
| [HellaSwag](https://noometry.com/benchmarks/hellaswag) | 79.2% | #18 of 29, top 63% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LAMBADA](https://noometry.com/benchmarks/lambada) | 75.2% | #5 of 9, top 56% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [PIQA](https://noometry.com/benchmarks/piqa) | 80.1% | #22 of 27, top 82% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [WinoGrande](https://noometry.com/benchmarks/winogrande) | 73% | #27 of 43, top 63% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Math

Llama 13b Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 838 | #285 of 285, top 100% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [GSM8K](https://noometry.com/benchmarks/gsm8k) | 20.6% | #34 of 38, top 90% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Knowledge

Llama 13b Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [ARC (AI2) Challenge](https://noometry.com/benchmarks/arc-challenge) | 52.7% | #27 of 39, top 70% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [BoolQ](https://noometry.com/benchmarks/boolq) | 78.7% | #16 of 23, top 70% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [MMLU](https://noometry.com/benchmarks/mmlu) | 47.7% | #70 of 81, top 87% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [OpenBookQA](https://noometry.com/benchmarks/openbookqa) | 56.4% | #16 of 19, top 85% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [TriviaQA](https://noometry.com/benchmarks/triviaqa) | 77.9% | #14 of 25, top 57% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Multimodal

Llama 13b Multimodal benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [ScienceQA](https://noometry.com/benchmarks/scienceqa) | 43.3% | #5 of 6, top 84% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Multilingual

Llama 13b Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 819 | #297 of 297, top 100% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Llama 13b Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 781 | #298 of 298, top 100% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Llama 13b Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 834 | #297 of 297, top 100% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 794 | #295 of 295, top 100% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 753 | #294 of 295, top 100% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

## Compare Llama 13b

-   [Llama 13b vs DeepSeek LLM 67B](https://noometry.com/compare/deepseek-llm-67b-vs-llama-13b)
-   [Llama 13b vs Llama 2-70B](https://noometry.com/compare/llama-13b-vs-llama-2-70b)
-   [Llama 13b vs Claude 2](https://noometry.com/compare/claude-2-vs-llama-13b)
-   [Llama 13b vs GPT-3.5-turbo](https://noometry.com/compare/gpt-3-5-turbo-vs-llama-13b)
-   [Llama 13b vs Claude 2.1](https://noometry.com/compare/claude-2-1-vs-llama-13b)
-   [Llama 13b vs Mistral 7B](https://noometry.com/compare/llama-13b-vs-mistral-7b)
-   [Llama 13b vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-llama-13b)
-   [Llama 13b vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-llama-13b)
-   [Llama 13b vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-llama-13b)
-   [Llama 13b vs Kimi K3](https://noometry.com/compare/kimi-k3-vs-llama-13b)
-   [Llama 13b vs Grok 4.6](https://noometry.com/compare/grok-4-6-vs-llama-13b)
-   [Llama 13b vs Qwen3.8 Max](https://noometry.com/compare/llama-13b-vs-qwen3-8-max)
-   [Llama 13b vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-llama-13b)
-   [Llama 13b vs DeepSeek V4 Pro](https://noometry.com/compare/deepseek-v4-pro-vs-llama-13b)

## Other Meta models

-   [Muse Spark 1.3](https://noometry.com/models/muse-spark-1-3)54.8
-   [Muse Spark](https://noometry.com/models/muse-spark)50.6
-   [Muse Spark 1.2](https://noometry.com/models/muse-spark-1-2)50.3
-   [Muse Spark 1.1](https://noometry.com/models/muse-spark-1-1)49.9
-   [Muse Glimmer](https://noometry.com/models/muse-glimmer)41.7
-   [Codellama 70b Instruct](https://noometry.com/models/codellama-70b-instruct)33.7
-   [Llama 4 Maverick](https://noometry.com/models/llama-4-maverick)30.9
-   [Codellama 34b Instruct](https://noometry.com/models/codellama-34b-instruct)30.8

## Frequently asked questions

### How good is Llama 13b?

Llama 13b by Meta ranks 348th of 354 ranked models on the Noometry Index as of October 2026, with a score of 24.4. Its strongest category is math, where it ranks 256th.

### Is Llama 13b open source?

Yes. Llama 13b's weights are downloadable; check the license for commercial terms.

### What are Llama 13b's strengths and weaknesses?

Relative to other ranked models, Llama 13b places best in math, reasoning, coding and lowest in writing & preference, instruction following, multilingual.

### What is Llama 13b best at?

Its best category is math, where it ranks 256th on Noometry.

### Cite this page

Noometry. (2026). Llama 13b benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/llama-13b

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/llama-13b.md).
