Meta, open weights
Codellama 70b Instruct
Codellama 70b Instruct by Meta ranks 237th of 354 ranked models on the Noometry Index as of October 2026, with a score of 33.7. Its strongest category is coding, where it ranks 193rd.
Last verified
Specifications
- Noometry rank
- #237 of 354
- Index score
- 33.7
- Evidence
- Confirmed 7 results
- Provider
Meta
- Released
- Unknown
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 37.6
- Reasoning 20.1
- Multilingual 24.8
- Instruction Following 51.9
- Writing & Preference 33.4
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 37.6 | #193 | 2 |
| Reasoning | 20.1 | #242 | 1 |
| Multilingual | 24.8 | #288 | 1 |
| Instruction Following | 51.9 | #293 | 1 |
| Writing & Preference | 33.4 | #277 | 1 |
Strengths and weaknesses
Categories where Codellama 70b Instruct places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multilingual | 24.8 | −22.6 | #288 of 297, top 97% |
| Instruction Following | 51.9 | −19.4 | #293 of 305, top 97% |
Closest competitors
The models ranked just above and below Codellama 70b Instruct. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Devstral Small 2505 | #233 | 34.3 | $0.15 | 88 | Compare |
| Qwen1.5-110B | #234 | 34.2 | — | — | Compare |
| o1-mini | #235 | 34.0 | — | — | Compare |
| Qwen3.5-9B | #236 | 33.8 | $0.11 | — | Compare |
| Qwen3 8B | #238 | 33.7 | $0.31 | — | Compare |
| Grok-2 (Dec 2024) | #239 | 33.7 | — | — | Compare |
| GPT-4.1 mini | #240 | 33.6 | $0.70 | 86 | Compare |
| GPT-5 Nano | #241 | 33.5 | $0.14 | 4 | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| BigCodeBench Instruct | 40.7% | #31 of 64, top 49% | BigCodeBench | 2023-08-25 | |
| BigCodeBench Complete | 44% | BigCodeBench | 2023-08-25 | ||
| BigCodeBench Complete | 49.6% | #36 of 66, top 55% | BigCodeBench | 2023-08-25 | |
| HumanEval+ | 65.9% | #23 of 45, top 52% | EvalPlus | ||
| HumanEval+ | 50.6% | EvalPlus |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1052 | #280 of 297, top 95% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 992 | #288 of 297, top 97% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1024 | #288 of 298, top 97% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1057 | #284 of 297, top 96% | LMArena | 2026-10-08 |
Compare Codellama 70b Instruct
- Codellama 70b Instruct vs Qwen3.5-9B
- Codellama 70b Instruct vs Qwen3 8B
- Codellama 70b Instruct vs o1-mini
- Codellama 70b Instruct vs Grok-2 (Dec 2024)
- Codellama 70b Instruct vs Qwen1.5-110B
- Codellama 70b Instruct vs GPT-4.1 mini
- Codellama 70b Instruct vs GPT-6 Astra
- Codellama 70b Instruct vs Claude Fable 5.1
- Codellama 70b Instruct vs Gemini 3.8 Flash
- Codellama 70b Instruct vs Kimi K3
- Codellama 70b Instruct vs Grok 4.6
- Codellama 70b Instruct vs Qwen3.8 Max
- Codellama 70b Instruct vs GLM-5.3
- Codellama 70b Instruct vs DeepSeek V4 Pro
Other Meta models
- Muse Spark 1.354.8
- Muse Spark50.6
- Muse Spark 1.250.3
- Muse Spark 1.149.9
- Muse Glimmer41.7
- Llama 4 Maverick30.9
- Codellama 34b Instruct30.8
- Llama 3.1-405B30.7
Frequently asked questions
How good is Codellama 70b Instruct?
Codellama 70b Instruct by Meta ranks 237th of 354 ranked models on the Noometry Index as of October 2026, with a score of 33.7. Its strongest category is coding, where it ranks 193rd.
Is Codellama 70b Instruct open source?
Yes. Codellama 70b Instruct's weights are downloadable; check the license for commercial terms.
What are Codellama 70b Instruct's strengths and weaknesses?
Relative to other ranked models, Codellama 70b Instruct places best in coding, reasoning and lowest in multilingual, instruction following.
What is Codellama 70b Instruct best at?
Its best category is coding, where it ranks 193rd on Noometry.