Meta, open weights
Llama 3.2 1B
Llama 3.2 1B by Meta ranks 354th of 354 ranked models on the Noometry Index as of October 2026, with a score of 20.1. Its strongest category is agentic & tool use, where it ranks 150th. API pricing starts at $0.027 per million input tokens and $0.20 per million output tokens, with a 60K-token context window.
Last verified
Specifications
- Noometry rank
- #354 of 354
- Index score
- 20.1
- Evidence
- Confirmed 22 results
- Provider
Meta
- Released
- September 24, 2024
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- 60K
- Max output
- 54K
- Input price
- $0.027 / M
- Output price
- $0.20 / M
- Blended price
- $0.0705 / M
- Output speed
- Not measured
- Value
- #21 of 219
- Knowledge cutoff
- Unknown
- Input
- text
- Hugging Face
- meta-llama/Llama-3.2-1B-Instruct
Category scores
Each category score combines every public result we have in that category.
- Coding 21.1
- Agentic & Tool Use 14.6
- Reasoning 16.2
- Math 10.4
- Knowledge 7.2
- Multilingual 23.8
- Instruction Following 52.4
- Long Context 31.9
- Writing & Preference 21.3
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 21.1 | #338 | 3 |
| Agentic & Tool Use | 14.6 | #150 | 2 |
| Reasoning | 16.2 | #308 | 2 |
| Math | 10.4 | #313 | 2 |
| Knowledge | 7.2 | #312 | 2 |
| Multilingual | 23.8 | #292 | 1 |
| Instruction Following | 52.4 | #290 | 1 |
| Long Context | 31.9 | #274 | 1 |
| Writing & Preference | 21.3 | #310 | 4 |
Strengths and weaknesses
Categories where Llama 3.2 1B places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Reasoning | 16.2 | −7.4 | #308 of 350, top 88% |
| Long Context | 31.9 | −9.0 | #274 of 296, top 93% |
| Instruction Following | 52.4 | −18.9 | #290 of 305, top 96% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Coding | 21.1 | −17.7 | #338 of 340, top 100% |
| Knowledge | 7.2 | −30.1 | #312 of 314, top 100% |
| Writing & Preference | 21.3 | −32.5 | #310 of 312, top 100% |
Closest competitors
The models ranked just above and below Llama 3.2 1B. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Claude 2 | #346 | 25.0 | — | — | Compare |
| DeepSeek LLM 67B | #347 | 24.9 | — | — | Compare |
| Llama 13b | #348 | 24.4 | — | — | Compare |
| Llama 2-70B | #349 | 24.4 | — | — | Compare |
| GPT-3.5-turbo | #350 | 23.2 | $0.75 | — | Compare |
| Mistral 7B | #351 | 23.0 | $0.25 | — | Compare |
| Llama 3.1-8B | #352 | 23.0 | $0.0575 | — | Compare |
| Gemma 3 1B | #353 | 21.1 | — | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| BigCodeBench Instruct | 8.2% | #63 of 64, top 99% | BigCodeBench | 2024-09-25 | |
| LMArena Coding | 1070 | #280 of 294, top 96% | LMArena | 2026-10-08 | |
| BigCodeBench Complete | 11.3% | #65 of 66, top 99% | BigCodeBench | 2024-09-25 |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Berkeley Function Calling Leaderboard | 10.8% | #47 of 49, top 96% | fc | Berkeley Function Calling Leaderboard | |
| BALROG | 6.6% | #35 of 35, top 100% | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Chess Puzzles | 0% | #121 of 129, top 94% | Epoch AI | 2026-08-30 | |
| LMArena Hard Prompts | 1044 | #283 of 297, top 96% | LMArena | 2026-10-08 | |
| Epoch Capabilities Index | 101.99 | #201 of 213, top 95% | Epoch AI | 2024-09-24 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| OTIS Mock AIME 2024-2025 | 0.6% | #171 of 173, top 99% | Epoch AI | 2026-08-30 | |
| LMArena Math | 1086 | #269 of 285, top 95% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 23.9% | #182 of 186, top 98% | Epoch AI | 2026-08-30 | |
| LMArena Expert | 1007 | #269 of 273, top 99% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 973 | #292 of 297, top 99% | LMArena | 2026-10-08 | |
| LMArena Chinese | 959 | #283 of 285, top 100% | LMArena | 2026-10-08 | |
| LMArena German | 1014 | #225 of 231, top 98% | LMArena | 2026-10-08 | |
| LMArena Russian | 941 | #282 of 283, top 100% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1031 | #286 of 298, top 96% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1050 | #279 of 291, top 96% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1055 | #286 of 297, top 97% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1033 | #284 of 295, top 97% | LMArena | 2026-10-08 | |
| EQ-Bench Creative Writing | 200 | #115 of 115, top 100% | EQ-Bench | ||
| LMArena Multi-Turn | 1030 | #281 of 295, top 96% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| openrouter | $0.027 | $0.20 | — | 2026-10-10 |
Compare Llama 3.2 1B
- Llama 3.2 1B vs Llama 3.1-405B
- Llama 3.2 1B vs Gemma 3 1B
- Llama 3.2 1B vs Llama 3.1-8B
- Llama 3.2 1B vs Mistral 7B
- Llama 3.2 1B vs GPT-3.5-turbo
- Llama 3.2 1B vs Llama 2-70B
- Llama 3.2 1B vs Llama 13b
- Llama 3.2 1B vs GPT-6 Astra
- Llama 3.2 1B vs Claude Fable 5.1
- Llama 3.2 1B vs Gemini 3.8 Flash
- Llama 3.2 1B vs Kimi K3
- Llama 3.2 1B vs Grok 4.6
- Llama 3.2 1B vs Qwen3.8 Max
- Llama 3.2 1B vs GLM-5.3
Other Meta models
- Muse Spark 1.354.8
- Muse Spark50.6
- Muse Spark 1.250.3
- Muse Spark 1.149.9
- Muse Glimmer41.7
- Codellama 70b Instruct33.7
- Llama 4 Maverick30.9
- Codellama 34b Instruct30.8
Frequently asked questions
How good is Llama 3.2 1B?
Llama 3.2 1B by Meta ranks 354th of 354 ranked models on the Noometry Index as of October 2026, with a score of 20.1. Its strongest category is agentic & tool use, where it ranks 150th. API pricing starts at $0.027 per million input tokens and $0.20 per million output tokens, with a 60K-token context window.
How much does Llama 3.2 1B cost?
Llama 3.2 1B costs $0.027 per million input tokens and $0.20 per million output tokens on openrouter.
What is Llama 3.2 1B's context window?
Llama 3.2 1B accepts up to 60K tokens of input and can write up to 54K tokens in one response.
Is Llama 3.2 1B open source?
Yes. Llama 3.2 1B's weights are downloadable from Hugging Face (meta-llama/Llama-3.2-1B-Instruct); check the license for commercial terms.
What are Llama 3.2 1B's strengths and weaknesses?
Relative to other ranked models, Llama 3.2 1B places best in reasoning, long context, instruction following and lowest in coding, knowledge, writing & preference.
What is Llama 3.2 1B best at?
Its best category is agentic & tool use, where it ranks 150th on Noometry.