Model comparison

Gemma 2 27B vs Nvidia Llama 3.3 Nemotron Super 49b v1.5

Nvidia Llama 3.3 Nemotron Super 49b v1.5 is the stronger model overall, scoring 40.3 to 29.4 on the Noometry Index.

Last verified . 12 shared benchmarks.

Gemma 2 27B Google

29.4

Rank #312 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Gemma 2 27B scores higher in 0 categories and Nvidia Llama 3.3 Nemotron Super 49b v1.5 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads 38.2 to 10.7.
  • Nvidia Llama 3.3 Nemotron Super 49b v1.5 is cheaper at $0.40 / $0.40 per million input/output tokens, against $0.65 / $0.65 for Gemma 2 27B.
  • Nvidia Llama 3.3 Nemotron Super 49b v1.5 accepts more context: 131K tokens versus 8K.

Side by side

Gemma 2 27B and Nvidia Llama 3.3 Nemotron Super 49b v1.5 specifications
Gemma 2 27BNvidia Llama 3.3 Nemotron Super 49b v1.5
ProviderGoogleNVIDIA
Noometry Index29.440.3
Released2024-06-242025-07-25
WeightsOpenOpen
Context window8K131K
Max output2K131K
Input $ / M tokens$0.65$0.40
Output $ / M tokens$0.65$0.40
Results tracked3412

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Gemma 2 27B: 34.1 (#246), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 39.8 (#154)

Coding benchmarks
BenchmarkGemma 2 27BNvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Coding12111355
BigCodeBench Instruct42.8%—
LiveBench Coding36%—
BigCodeBench Complete52.5%—

Reasoning Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Gemma 2 27B: 15.3 (#315), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 26.8 (#128)

Reasoning benchmarks
BenchmarkGemma 2 27BNvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Hard Prompts11981336
LiveBench Reasoning28.1%—
DTBench48%—
LiveBench Data Analysis47.9%—
LMCA7.1%—
Epoch Capabilities Index122.08—
LiveBench38.2%—

Math Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Gemma 2 27B: 10.7 (#311), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 38.2 (#141)

Math benchmarks
BenchmarkGemma 2 27BNvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Math12121392
OTIS Mock AIME 2024-20251.4%—
LiveBench Math26.5%—
MATH Level 527.9%—

Knowledge Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Gemma 2 27B: 19.0 (#280), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 36.7 (#165)

Knowledge benchmarks
BenchmarkGemma 2 27BNvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Expert11721330
GPQA Diamond36.5%—
Confabulations27.1%—
MMLU75.7%—

Multilingual Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Gemma 2 27B: 38.6 (#226), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 45.5 (#168)

Multilingual benchmarks
BenchmarkGemma 2 27BNvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Non-English12171316
LMArena Japanese11751300
LMArena Russian12341332
LMArena Chinese1221—
LMArena French1247—
LMArena German1209—
LMArena Korean1174—
LMArena Spanish1228—

Instruction Following Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Gemma 2 27B: 60.5 (#249), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 68.6 (#188)

Instruction Following benchmarks
BenchmarkGemma 2 27BNvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Instruction Following12061299
LiveBench Instruction Following58.1%—

Long Context Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Gemma 2 27B: 37.3 (#218), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 40.0 (#164)

Long Context benchmarks
BenchmarkGemma 2 27BNvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Longer Query12311315

Writing & Preference Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads

Gemma 2 27B: 44.2 (#225), Nvidia Llama 3.3 Nemotron Super 49b v1.5: 53.1 (#159)

Writing & Preference benchmarks
BenchmarkGemma 2 27BNvidia Llama 3.3 Nemotron Super 49b v1.5
LMArena Text12311338
LMArena Creative Writing12411307
LMArena Multi-Turn12241334
LiveBench Language32.6%—

Frequently asked questions

Is Gemma 2 27B better than Nvidia Llama 3.3 Nemotron Super 49b v1.5?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 is the stronger model overall, scoring 40.3 to 29.4 on the Noometry Index.

Which is cheaper, Gemma 2 27B or Nvidia Llama 3.3 Nemotron Super 49b v1.5?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 is cheaper. It lists at $0.40 per million input tokens and $0.40 per million output tokens; Gemma 2 27B lists at $0.65 and $0.65.

Is Gemma 2 27B or Nvidia Llama 3.3 Nemotron Super 49b v1.5 better for coding?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 scores higher on coding benchmarks: 39.8 versus 34.1 in the Noometry coding category.

Which has the bigger context window?

Nvidia Llama 3.3 Nemotron Super 49b v1.5 does, with 131K tokens against 8K.

How many benchmarks do Gemma 2 27B and Nvidia Llama 3.3 Nemotron Super 49b v1.5 share?

12 benchmarks have published results for both models. Gemma 2 27B has 34 scored results on Noometry and Nvidia Llama 3.3 Nemotron Super 49b v1.5 has 12.

Related comparisons

Go deeper