Model comparison

Gemma 2 9B vs Hunyuan Turbos 20250226

Hunyuan Turbos 20250226 is the stronger model overall, scoring 41.3 to 25.9 on the Noometry Index.

Last verified . 16 shared benchmarks.

Gemma 2 9B Google

25.9

Rank #341 Confirmed

Hunyuan Turbos 20250226 Tencent

41.3

Rank #139 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Gemma 2 9B scores higher in 0 categories and Hunyuan Turbos 20250226 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Hunyuan Turbos 20250226 leads 37.5 to 9.9.
  • Gemma 2 9B has downloadable open weights; the other is API-only.

Side by side

Gemma 2 9B and Hunyuan Turbos 20250226 specifications
Gemma 2 9BHunyuan Turbos 20250226
ProviderGoogleTencent
Noometry Index25.941.3
Released2024-06-24—
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked3516

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Turbos 20250226 leads

Gemma 2 9B: 29.4 (#304), Hunyuan Turbos 20250226: 40.0 (#152)

Coding benchmarks
BenchmarkGemma 2 9BHunyuan Turbos 20250226
LMArena Coding11731361
BigCodeBench Instruct34.7%—
LiveBench Coding22.5%—
BigCodeBench Complete40.6%—

Reasoning Hunyuan Turbos 20250226 leads

Gemma 2 9B: 15.9 (#309), Hunyuan Turbos 20250226: 27.8 (#113)

Reasoning benchmarks
BenchmarkGemma 2 9BHunyuan Turbos 20250226
LMArena Hard Prompts11711374
LiveBench Reasoning15.2%—
LiveBench Data Analysis36.4%—
Epoch Capabilities Index119.83—
LiveBench28.7%—
PIQA83.7%—

Math Hunyuan Turbos 20250226 leads

Gemma 2 9B: 9.9 (#318), Hunyuan Turbos 20250226: 37.5 (#154)

Math benchmarks
BenchmarkGemma 2 9BHunyuan Turbos 20250226
LMArena Math11831359
OTIS Mock AIME 2024-20250.6%—
LiveBench Math19.8%—
MATH Level 521%—
GSM8K84.9%—

Knowledge Hunyuan Turbos 20250226 leads

Gemma 2 9B: 9.7 (#305), Hunyuan Turbos 20250226: 37.0 (#161)

Knowledge benchmarks
BenchmarkGemma 2 9BHunyuan Turbos 20250226
LMArena Expert11471339
GPQA Diamond27.5%—
BoolQ85.7%—
MMLU72.1%—

Multilingual Hunyuan Turbos 20250226 leads

Gemma 2 9B: 36.6 (#238), Hunyuan Turbos 20250226: 48.9 (#136)

Multilingual benchmarks
BenchmarkGemma 2 9BHunyuan Turbos 20250226
LMArena Non-English11881363
LMArena Chinese11851417
LMArena French11901391
LMArena German11861355
LMArena Japanese11441342
LMArena Korean11371351
LMArena Russian12001368
LMArena Spanish1200—

Instruction Following Hunyuan Turbos 20250226 leads

Gemma 2 9B: 57.6 (#269), Hunyuan Turbos 20250226: 71.0 (#158)

Instruction Following benchmarks
BenchmarkGemma 2 9BHunyuan Turbos 20250226
LMArena Instruction Following11781344
LiveBench Instruction Following52.6%—

Long Context Hunyuan Turbos 20250226 leads

Gemma 2 9B: 36.3 (#233), Hunyuan Turbos 20250226: 41.6 (#136)

Long Context benchmarks
BenchmarkGemma 2 9BHunyuan Turbos 20250226
LMArena Longer Query11971366

Writing & Preference Hunyuan Turbos 20250226 leads

Gemma 2 9B: 32.1 (#281), Hunyuan Turbos 20250226: 57.4 (#128)

Writing & Preference benchmarks
BenchmarkGemma 2 9BHunyuan Turbos 20250226
LMArena Text12071377
LMArena Creative Writing12061359
LMArena Multi-Turn11931387
EQ-Bench Creative Writing841—
LiveBench Language25.5%—

Frequently asked questions

Is Gemma 2 9B better than Hunyuan Turbos 20250226?

Hunyuan Turbos 20250226 is the stronger model overall, scoring 41.3 to 25.9 on the Noometry Index.

Is Gemma 2 9B or Hunyuan Turbos 20250226 better for coding?

Hunyuan Turbos 20250226 scores higher on coding benchmarks: 40.0 versus 29.4 in the Noometry coding category.

How many benchmarks do Gemma 2 9B and Hunyuan Turbos 20250226 share?

16 benchmarks have published results for both models. Gemma 2 9B has 35 scored results on Noometry and Hunyuan Turbos 20250226 has 16.

Related comparisons

Go deeper