Meta, open weights

Llama 2-7B

Llama 2-7B by Meta ranks 317th of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.1. Its strongest category is math, where it ranks 233rd.

Last verified

Specifications

Noometry rank
#317 of 354
Index score
29.1
Evidence
Confirmed 29 results
Provider
Meta
Released
July 18, 2023
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Llama 2-7B category scores
  1. Coding 29.2
  2. Reasoning 15.7
  3. Math 30.7
  4. Knowledge 28.2
  5. Multilingual 23.8
  6. Instruction Following 50.8
  7. Long Context 30.4
  8. Writing & Preference 28.0
Llama 2-7B category ranks
CategoryScoreRankResults
Coding29.2#3071
Reasoning15.7#3122
Math30.7#2331
Knowledge28.2#2481
Multilingual23.8#2931
Instruction Following50.8#2981
Long Context30.4#2871
Writing & Preference28.0#2983

Strengths and weaknesses

Categories where Llama 2-7B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Llama 2-7B: strongest categories
CategoryScorevs medianRank
Math30.7−5.9#233 of 327, top 72%
Knowledge28.2−9.1#248 of 314, top 79%
Reasoning15.7−7.9#312 of 350, top 90%

Weakest categories

Llama 2-7B: weakest categories
CategoryScorevs medianRank
Multilingual23.8−23.7#293 of 297, top 99%
Instruction Following50.8−20.5#298 of 305, top 98%
Long Context30.4−10.5#287 of 296, top 97%

Closest competitors

The models ranked just above and below Llama 2-7B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Llama 2-7B
ModelRankScoreBlended $/MSpeed
Gemma 1.1 2b IT#31329.3——Compare
Phi 3 Small 8k Instruct#31429.3——Compare
Claude 3.5 Haiku#31529.2——Compare
GPT-4#31629.1$37.50—Compare
Granite 4.0 Micro#31829.0$0.0408—Compare
Claude 3 Sonnet#31929.0——Compare
Qwen2.5 7B Instruct#32029.0$0.31—Compare
Llama 3.2 3B#32128.9$0.12—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Llama 2-7B Coding benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Coding1002#290 of 294, top 99%LMArena2026-10-08

Reasoning

Llama 2-7B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Chess Puzzles0%#119 of 129, top 93%Epoch AI2026-08-30
LMArena Hard Prompts1009#289 of 297, top 98%LMArena2026-10-08
BIG-Bench Hard39.2%#22 of 27, top 82%Epoch AI
Epoch Capabilities Index99.06#204 of 213, top 96%Epoch AI2023-07-18
HellaSwag77.2%#20 of 29, top 69%Epoch AI
LAMBADA73.3%#7 of 9, top 78%Epoch AI
PIQA78.8%#24 of 27, top 89%Epoch AI
WinoGrande69.2%#32 of 43, top 75%Epoch AI

Math

Llama 2-7B Math benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Math1042#279 of 285, top 98%LMArena2026-10-08
GSM8K16.7%#36 of 38, top 95%Epoch AI

Knowledge

Llama 2-7B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Expert1036#266 of 273, top 98%LMArena2026-10-08
ARC (AI2) Challenge45.9%#31 of 39, top 80%Epoch AI
BoolQ77.9%#18 of 23, top 79%Epoch AI
MMLU45.8%#71 of 81, top 88%Epoch AI
OpenBookQA58.6%#12 of 19, top 64%Epoch AI
TriviaQA73.7%#17 of 25, top 68%Epoch AI

Multimodal

Llama 2-7B Multimodal benchmark results
BenchmarkScorePositionSettingSourceDate
ScienceQA43.1%#6 of 6, top 100%Epoch AI

Multilingual

Llama 2-7B Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English973#293 of 297, top 99%LMArena2026-10-08
LMArena Chinese973#282 of 285, top 99%LMArena2026-10-08
LMArena French970#223 of 223, top 100%LMArena2026-10-08
LMArena German978#229 of 231, top 100%LMArena2026-10-08
LMArena Russian995#276 of 283, top 98%LMArena2026-10-08
LMArena Spanish1007#226 of 226, top 100%LMArena2026-10-08

Instruction Following

Llama 2-7B Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1006#292 of 298, top 98%LMArena2026-10-08

Long Context

Llama 2-7B Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query999#287 of 291, top 99%LMArena2026-10-08

Writing & Preference

Llama 2-7B Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1053#288 of 297, top 97%LMArena2026-10-08
LMArena Creative Writing1033#283 of 295, top 96%LMArena2026-10-08
LMArena Multi-Turn1029#282 of 295, top 96%LMArena2026-10-08

Compare Llama 2-7B

Other Meta models

Frequently asked questions

How good is Llama 2-7B?

Llama 2-7B by Meta ranks 317th of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.1. Its strongest category is math, where it ranks 233rd.

Is Llama 2-7B open source?

Yes. Llama 2-7B's weights are downloadable; check the license for commercial terms.

What are Llama 2-7B's strengths and weaknesses?

Relative to other ranked models, Llama 2-7B places best in math, knowledge, reasoning and lowest in multilingual, instruction following, long context.

What is Llama 2-7B best at?

Its best category is math, where it ranks 233rd on Noometry.