Model comparison

INTELLECT-1 vs Llama 2-34B

Neither INTELLECT-1 nor Llama 2-34B has enough public benchmark results to be ranked yet; the rows below show what has been published.

Last verified . 6 shared benchmarks.

Llama 2-34B Meta

—

Unranked

Summary

  • They share 6 benchmarks with published results for both.

Side by side

INTELLECT-1 and Llama 2-34B specifications
INTELLECT-1Llama 2-34B
ProviderHugging FaceMeta
Noometry Index——
Released2024-11-292023-07-18
WeightsProprietaryProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked710

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Reasoning Not comparable

INTELLECT-1: —, Llama 2-34B: —

Reasoning benchmarks
BenchmarkINTELLECT-1Llama 2-34B
BIG-Bench Hard34.9%44.1%
Epoch Capabilities Index100.86105.2
WinoGrande65.8%76.7%
HellaSwag71.4%—
PIQA—81.9%

Math Not comparable

INTELLECT-1: —, Llama 2-34B: —

Math benchmarks
BenchmarkINTELLECT-1Llama 2-34B
GSM8K38.6%42.2%

Knowledge Not comparable

INTELLECT-1: —, Llama 2-34B: —

Knowledge benchmarks
BenchmarkINTELLECT-1Llama 2-34B
ARC (AI2) Challenge54.5%54.5%
MMLU49.9%62.6%
BoolQ—83.7%
OpenBookQA—58.2%
TriviaQA—84.6%

Frequently asked questions

Is INTELLECT-1 better than Llama 2-34B?

Neither INTELLECT-1 nor Llama 2-34B has enough public benchmark results to be ranked yet; the rows below show what has been published.

How many benchmarks do INTELLECT-1 and Llama 2-34B share?

6 benchmarks have published results for both models. INTELLECT-1 has 7 scored results on Noometry and Llama 2-34B has 10.

Related comparisons

Go deeper