Model comparison

Gemma 7B vs Hunyuan T1 20250711

Hunyuan T1 20250711 is the stronger model overall, scoring 42.5 to 30.0 on the Noometry Index.

Last verified . 12 shared benchmarks.

Gemma 7B Google

30.0

Rank #299 Confirmed

Hunyuan T1 20250711 Tencent

42.5

Rank #114 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Gemma 7B scores higher in 0 categories and Hunyuan T1 20250711 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Hunyuan T1 20250711 leads 59.5 to 27.1.
  • Gemma 7B has downloadable open weights; the other is API-only.

Side by side

Gemma 7B and Hunyuan T1 20250711 specifications
Gemma 7BHunyuan T1 20250711
ProviderGoogleTencent
Noometry Index30.042.5
Released2024-02-21—
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked2713

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan T1 20250711 leads

Gemma 7B: 30.5 (#294), Hunyuan T1 20250711: 40.9 (#129)

Coding benchmarks
BenchmarkGemma 7BHunyuan T1 20250711
LMArena Coding10481390
HumanEval+28.7%—
MBPP+43.4%—

Reasoning Hunyuan T1 20250711 leads

Gemma 7B: 19.9 (#249), Hunyuan T1 20250711: 28.5 (#103)

Reasoning benchmarks
BenchmarkGemma 7BHunyuan T1 20250711
LMArena Hard Prompts10421399
Adversarial NLI48.7%—
BIG-Bench Hard55.1%—
Epoch Capabilities Index111.99—
HellaSwag82.2%—
PIQA81.2%—
WinoGrande79%—

Math Hunyuan T1 20250711 leads

Gemma 7B: 31.2 (#228), Hunyuan T1 20250711: 38.7 (#130)

Math benchmarks
BenchmarkGemma 7BHunyuan T1 20250711
LMArena Math10661414
GSM8K46.4%—

Knowledge Hunyuan T1 20250711 leads

Gemma 7B: 27.3 (#252), Hunyuan T1 20250711: 38.8 (#141)

Knowledge benchmarks
BenchmarkGemma 7BHunyuan T1 20250711
LMArena Expert10011395
ARC (AI2) Challenge78.3%—
BoolQ83.2%—
MMLU66.1%—
OpenBookQA78.6%—
TriviaQA72.3%—

Multilingual Hunyuan T1 20250711 leads

Gemma 7B: 25.1 (#287), Hunyuan T1 20250711: 51.2 (#112)

Multilingual benchmarks
BenchmarkGemma 7BHunyuan T1 20250711
LMArena Non-English9991395
LMArena Chinese10351425
LMArena Russian9931385
LMArena French1025—
LMArena Korean—1406

Instruction Following Hunyuan T1 20250711 leads

Gemma 7B: 51.5 (#295), Hunyuan T1 20250711: 72.6 (#138)

Instruction Following benchmarks
BenchmarkGemma 7BHunyuan T1 20250711
LMArena Instruction Following10171374

Long Context Hunyuan T1 20250711 leads

Gemma 7B: 31.1 (#282), Hunyuan T1 20250711: 42.2 (#128)

Long Context benchmarks
BenchmarkGemma 7BHunyuan T1 20250711
LMArena Longer Query10221384

Writing & Preference Hunyuan T1 20250711 leads

Gemma 7B: 27.1 (#302), Hunyuan T1 20250711: 59.5 (#109)

Writing & Preference benchmarks
BenchmarkGemma 7BHunyuan T1 20250711
LMArena Text10561401
LMArena Creative Writing10241392
LMArena Multi-Turn9631393

Frequently asked questions

Is Gemma 7B better than Hunyuan T1 20250711?

Hunyuan T1 20250711 is the stronger model overall, scoring 42.5 to 30.0 on the Noometry Index.

Is Gemma 7B or Hunyuan T1 20250711 better for coding?

Hunyuan T1 20250711 scores higher on coding benchmarks: 40.9 versus 30.5 in the Noometry coding category.

How many benchmarks do Gemma 7B and Hunyuan T1 20250711 share?

12 benchmarks have published results for both models. Gemma 7B has 27 scored results on Noometry and Hunyuan T1 20250711 has 13.

Related comparisons

Go deeper