Model comparison

Llama 2-34B vs Mercury 2.5

Mercury 2.5 has enough public results to be ranked (#242); Llama 2-34B does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

Llama 2-34B Meta

—

Unranked

Mercury 2.5 Inception

33.5

Rank #242 Reported

Side by side

Llama 2-34B and Mercury 2.5 specifications
Llama 2-34BMercury 2.5
ProviderMetaInception
Noometry Index—33.5
Released2023-07-182026-09-08
WeightsProprietaryProprietary
Context window—260K
Max output—66K
Input $ / M tokens—$0.04
Output $ / M tokens—$0.15
Results tracked104

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Llama 2-34B: —, Mercury 2.5: 39.5 (#156)

Coding benchmarks
BenchmarkLlama 2-34BMercury 2.5
SciCode—38.5%
ALE-Bench—301.65

Reasoning Not comparable

Llama 2-34B: —, Mercury 2.5: 22.4 (#193)

Reasoning benchmarks
BenchmarkLlama 2-34BMercury 2.5
CritPt—0%
BIG-Bench Hard44.1%—
Epoch Capabilities Index105.2—
PIQA81.9%—
WinoGrande76.7%—

Math Not comparable

Llama 2-34B: —, Mercury 2.5: 23.3 (#272)

Math benchmarks
BenchmarkLlama 2-34BMercury 2.5
ProofBench—3%
GSM8K42.2%—

Knowledge Not comparable

Llama 2-34B: —, Mercury 2.5: —

Knowledge benchmarks
BenchmarkLlama 2-34BMercury 2.5
ARC (AI2) Challenge54.5%—
BoolQ83.7%—
MMLU62.6%—
OpenBookQA58.2%—
TriviaQA84.6%—

Frequently asked questions

Is Llama 2-34B better than Mercury 2.5?

Mercury 2.5 has enough public results to be ranked (#242); Llama 2-34B does not yet, so treat this comparison as directional.

How many benchmarks do Llama 2-34B and Mercury 2.5 share?

0 benchmarks have published results for both models. Llama 2-34B has 10 scored results on Noometry and Mercury 2.5 has 4.

Related comparisons

Go deeper