Model comparison

Gemma 2B vs Hunyuan Standard 2025 02 10

Hunyuan Standard 2025 02 10 is the stronger model overall, scoring 37.9 to 29.6 on the Noometry Index.

Last verified . 11 shared benchmarks.

Gemma 2B Google

29.6

Rank #307 Confirmed

Hunyuan Standard 2025 02 10 Tencent

37.9

Rank #193 Confirmed

Summary

  • They share 11 benchmarks with published results for both. Gemma 2B scores higher in 0 categories and Hunyuan Standard 2025 02 10 in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Hunyuan Standard 2025 02 10 leads 47.2 to 24.0.
  • Gemma 2B has downloadable open weights; the other is API-only.

Side by side

Gemma 2B and Hunyuan Standard 2025 02 10 specifications
Gemma 2BHunyuan Standard 2025 02 10
ProviderGoogleTencent
Noometry Index29.637.9
Released2024-02-21—
WeightsOpenProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked2312

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Standard 2025 02 10 leads

Gemma 2B: 29.4 (#305), Hunyuan Standard 2025 02 10: 37.1 (#197)

Coding benchmarks
BenchmarkGemma 2BHunyuan Standard 2025 02 10
LMArena Coding10101270
HumanEval+20.7%—
MBPP+34.1%—

Reasoning Hunyuan Standard 2025 02 10 leads

Gemma 2B: 18.8 (#275), Hunyuan Standard 2025 02 10: 25.0 (#154)

Reasoning benchmarks
BenchmarkGemma 2BHunyuan Standard 2025 02 10
LMArena Hard Prompts9891264
BIG-Bench Hard35.2%—
Epoch Capabilities Index94.2—
HellaSwag71.4%—
PIQA77.3%—
WinoGrande65.4%—

Math Hunyuan Standard 2025 02 10 leads

Gemma 2B: 30.0 (#239), Hunyuan Standard 2025 02 10: 35.6 (#179)

Math benchmarks
BenchmarkGemma 2BHunyuan Standard 2025 02 10
LMArena Math10091274
GSM8K17.7%—

Knowledge Not comparable

Gemma 2B: —, Hunyuan Standard 2025 02 10: 34.2 (#197)

Knowledge benchmarks
BenchmarkGemma 2BHunyuan Standard 2025 02 10
LMArena Expert—1248
ARC (AI2) Challenge42.1%—
BoolQ69.4%—
MMLU42.3%—
TriviaQA53.2%—

Multilingual Hunyuan Standard 2025 02 10 leads

Gemma 2B: 23.0 (#294), Hunyuan Standard 2025 02 10: 41.7 (#203)

Multilingual benchmarks
BenchmarkGemma 2BHunyuan Standard 2025 02 10
LMArena Non-English9581262
LMArena Chinese9861319
LMArena Russian9371258

Instruction Following Hunyuan Standard 2025 02 10 leads

Gemma 2B: 48.5 (#302), Hunyuan Standard 2025 02 10: 65.5 (#219)

Instruction Following benchmarks
BenchmarkGemma 2BHunyuan Standard 2025 02 10
LMArena Instruction Following9701245

Long Context Hunyuan Standard 2025 02 10 leads

Gemma 2B: 29.9 (#291), Hunyuan Standard 2025 02 10: 39.5 (#173)

Long Context benchmarks
BenchmarkGemma 2BHunyuan Standard 2025 02 10
LMArena Longer Query9811301

Writing & Preference Hunyuan Standard 2025 02 10 leads

Gemma 2B: 24.0 (#308), Hunyuan Standard 2025 02 10: 47.2 (#214)

Writing & Preference benchmarks
BenchmarkGemma 2BHunyuan Standard 2025 02 10
LMArena Text10021274
LMArena Creative Writing9871242
LMArena Multi-Turn9451275

Frequently asked questions

Is Gemma 2B better than Hunyuan Standard 2025 02 10?

Hunyuan Standard 2025 02 10 is the stronger model overall, scoring 37.9 to 29.6 on the Noometry Index.

Is Gemma 2B or Hunyuan Standard 2025 02 10 better for coding?

Hunyuan Standard 2025 02 10 scores higher on coding benchmarks: 37.1 versus 29.4 in the Noometry coding category.

How many benchmarks do Gemma 2B and Hunyuan Standard 2025 02 10 share?

11 benchmarks have published results for both models. Gemma 2B has 23 scored results on Noometry and Hunyuan Standard 2025 02 10 has 12.

Related comparisons

Go deeper