Model comparison

Hunyuan Standard 2025 02 10 vs Mercury 2

Mercury 2 is the stronger model overall, scoring 39.1 to 37.9 on the Noometry Index.

Last verified . 11 shared benchmarks.

Hunyuan Standard 2025 02 10 Tencent

37.9

Rank #193 Confirmed

Mercury 2 Inception

39.1

Rank #175 Confirmed

Summary

  • They share 11 benchmarks with published results for both. Hunyuan Standard 2025 02 10 scores higher in 2 categories and Mercury 2 in 5 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Mercury 2 leads 53.8 to 47.2.

Side by side

Hunyuan Standard 2025 02 10 and Mercury 2 specifications
Hunyuan Standard 2025 02 10Mercury 2
ProviderTencentInception
Noometry Index37.939.1
Released—2026-02-20
WeightsProprietaryProprietary
Context window—128K
Max output—50K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.75
Results tracked1217

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Standard 2025 02 10 leads

Hunyuan Standard 2025 02 10: 37.1 (#197), Mercury 2: 33.5 (#255)

Coding benchmarks
BenchmarkHunyuan Standard 2025 02 10Mercury 2
LMArena Coding12701391
LMArena WebDev—1171
SciCode—38.7%
WeirdML—43.2%
ALE-Bench—785.58

Reasoning Hunyuan Standard 2025 02 10 leads

Hunyuan Standard 2025 02 10: 25.0 (#154), Mercury 2: 23.8 (#170)

Reasoning benchmarks
BenchmarkHunyuan Standard 2025 02 10Mercury 2
LMArena Hard Prompts12641362
CritPt—0.8%

Math Not comparable

Hunyuan Standard 2025 02 10: 35.6 (#179), Mercury 2: —

Math benchmarks
BenchmarkHunyuan Standard 2025 02 10Mercury 2
LMArena Math1274—

Knowledge Mercury 2 leads

Hunyuan Standard 2025 02 10: 34.2 (#197), Mercury 2: 36.2 (#172)

Knowledge benchmarks
BenchmarkHunyuan Standard 2025 02 10Mercury 2
LMArena Expert12481358
Vectara Hallucination Rate—12.3%

Multilingual Mercury 2 leads

Hunyuan Standard 2025 02 10: 41.7 (#203), Mercury 2: 46.6 (#157)

Multilingual benchmarks
BenchmarkHunyuan Standard 2025 02 10Mercury 2
LMArena Non-English12621331
LMArena Chinese13191417
LMArena Russian12581304

Instruction Following Mercury 2 leads

Hunyuan Standard 2025 02 10: 65.5 (#219), Mercury 2: 70.2 (#165)

Instruction Following benchmarks
BenchmarkHunyuan Standard 2025 02 10Mercury 2
LMArena Instruction Following12451329

Long Context Too close to call

Hunyuan Standard 2025 02 10: 39.5 (#173), Mercury 2: 40.5 (#154)

Long Context benchmarks
BenchmarkHunyuan Standard 2025 02 10Mercury 2
LMArena Longer Query13011330

Writing & Preference Mercury 2 leads

Hunyuan Standard 2025 02 10: 47.2 (#214), Mercury 2: 53.8 (#155)

Writing & Preference benchmarks
BenchmarkHunyuan Standard 2025 02 10Mercury 2
LMArena Text12741355
LMArena Creative Writing12421289
LMArena Multi-Turn12751358

Frequently asked questions

Is Hunyuan Standard 2025 02 10 better than Mercury 2?

Mercury 2 is the stronger model overall, scoring 39.1 to 37.9 on the Noometry Index.

Is Hunyuan Standard 2025 02 10 or Mercury 2 better for coding?

Hunyuan Standard 2025 02 10 scores higher on coding benchmarks: 37.1 versus 33.5 in the Noometry coding category.

How many benchmarks do Hunyuan Standard 2025 02 10 and Mercury 2 share?

11 benchmarks have published results for both models. Hunyuan Standard 2025 02 10 has 12 scored results on Noometry and Mercury 2 has 17.

Related comparisons

Go deeper