Model comparison

Hunyuan Large 2025 02 10 vs Trinity Large Thinking

Hunyuan Large 2025 02 10 and Trinity Large Thinking score almost the same on the Noometry Index (38.6 vs 38.6), so choose on price, context window or the category you care about most.

Last verified . 12 shared benchmarks.

Hunyuan Large 2025 02 10 Tencent

38.6

Rank #184 Confirmed

Trinity Large Thinking Arcee AI

38.6

Rank #185 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Hunyuan Large 2025 02 10 scores higher in 2 categories and Trinity Large Thinking in 6 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Hunyuan Large 2025 02 10 leads 25.5 to 16.9.
  • Trinity Large Thinking has downloadable open weights; the other is API-only.

Side by side

Hunyuan Large 2025 02 10 and Trinity Large Thinking specifications
Hunyuan Large 2025 02 10Trinity Large Thinking
ProviderTencentArcee AI
Noometry Index38.638.6
Released—2026-04-01
WeightsProprietaryOpen
Context window—262K
Max output—80K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.80
Results tracked1224

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 38.2 (#181), Trinity Large Thinking: 34.1 (#244)

Coding benchmarks
BenchmarkHunyuan Large 2025 02 10Trinity Large Thinking
LMArena Coding13071381
LMArena WebDev—1238
SciCode—36.1%

Reasoning Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 25.5 (#148), Trinity Large Thinking: 16.9 (#298)

Reasoning benchmarks
BenchmarkHunyuan Large 2025 02 10Trinity Large Thinking
LMArena Hard Prompts12861350
NYT Connections (extended)—16.5%
CritPt—0.9%
Thematic Generalization—41.6%
Surface Evolver Bench—15.6%

Math Trinity Large Thinking leads

Hunyuan Large 2025 02 10: 35.8 (#178), Trinity Large Thinking: 37.6 (#149)

Math benchmarks
BenchmarkHunyuan Large 2025 02 10Trinity Large Thinking
LMArena Math12811366

Knowledge Trinity Large Thinking leads

Hunyuan Large 2025 02 10: 35.1 (#188), Trinity Large Thinking: 40.9 (#113)

Knowledge benchmarks
BenchmarkHunyuan Large 2025 02 10Trinity Large Thinking
LMArena Expert12761360
Vectara Hallucination Rate—6.9%

Multilingual Trinity Large Thinking leads

Hunyuan Large 2025 02 10: 42.0 (#200), Trinity Large Thinking: 46.2 (#160)

Multilingual benchmarks
BenchmarkHunyuan Large 2025 02 10Trinity Large Thinking
LMArena Non-English12651325
LMArena Chinese13461373
LMArena Russian12661337
LMArena French—1374
LMArena German—1356
LMArena Japanese—1311
LMArena Korean—1306
LMArena Spanish—1357

Instruction Following Trinity Large Thinking leads

Hunyuan Large 2025 02 10: 67.3 (#197), Trinity Large Thinking: 70.5 (#162)

Instruction Following benchmarks
BenchmarkHunyuan Large 2025 02 10Trinity Large Thinking
LMArena Instruction Following12771334

Long Context Too close to call

Hunyuan Large 2025 02 10: 40.8 (#149), Trinity Large Thinking: 41.3 (#144)

Long Context benchmarks
BenchmarkHunyuan Large 2025 02 10Trinity Large Thinking
LMArena Longer Query13411355

Writing & Preference Trinity Large Thinking leads

Hunyuan Large 2025 02 10: 48.7 (#197), Trinity Large Thinking: 53.8 (#158)

Writing & Preference benchmarks
BenchmarkHunyuan Large 2025 02 10Trinity Large Thinking
LMArena Text12881340
LMArena Creative Writing12641320
LMArena Multi-Turn12841342

Frequently asked questions

Is Hunyuan Large 2025 02 10 better than Trinity Large Thinking?

Hunyuan Large 2025 02 10 and Trinity Large Thinking score almost the same on the Noometry Index (38.6 vs 38.6), so choose on price, context window or the category you care about most.

Is Hunyuan Large 2025 02 10 or Trinity Large Thinking better for coding?

Hunyuan Large 2025 02 10 scores higher on coding benchmarks: 38.2 versus 34.1 in the Noometry coding category.

How many benchmarks do Hunyuan Large 2025 02 10 and Trinity Large Thinking share?

12 benchmarks have published results for both models. Hunyuan Large 2025 02 10 has 12 scored results on Noometry and Trinity Large Thinking has 24.

Related comparisons

Go deeper