Model comparison

Hunyuan Turbos 20250226 vs Trinity Large Thinking

Hunyuan Turbos 20250226 is the stronger model overall, scoring 41.3 to 38.6 on the Noometry Index.

Last verified . 16 shared benchmarks.

Hunyuan Turbos 20250226 Tencent

41.3

Rank #139 Confirmed

Trinity Large Thinking Arcee AI

38.6

Rank #185 Confirmed

Summary

  • They share 16 benchmarks with published results for both. Hunyuan Turbos 20250226 scores higher in 6 categories and Trinity Large Thinking in 2 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Hunyuan Turbos 20250226 leads 27.8 to 16.9.
  • Trinity Large Thinking has downloadable open weights; the other is API-only.

Side by side

Hunyuan Turbos 20250226 and Trinity Large Thinking specifications
Hunyuan Turbos 20250226Trinity Large Thinking
ProviderTencentArcee AI
Noometry Index41.338.6
Released—2026-04-01
WeightsProprietaryOpen
Context window—262K
Max output—80K
Input $ / M tokens—$0.25
Output $ / M tokens—$0.80
Results tracked1624

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 40.0 (#152), Trinity Large Thinking: 34.1 (#244)

Coding benchmarks
BenchmarkHunyuan Turbos 20250226Trinity Large Thinking
LMArena Coding13611381
LMArena WebDev—1238
SciCode—36.1%

Reasoning Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 27.8 (#113), Trinity Large Thinking: 16.9 (#298)

Reasoning benchmarks
BenchmarkHunyuan Turbos 20250226Trinity Large Thinking
LMArena Hard Prompts13741350
NYT Connections (extended)—16.5%
CritPt—0.9%
Thematic Generalization—41.6%
Surface Evolver Bench—15.6%

Math Too close to call

Hunyuan Turbos 20250226: 37.5 (#154), Trinity Large Thinking: 37.6 (#149)

Math benchmarks
BenchmarkHunyuan Turbos 20250226Trinity Large Thinking
LMArena Math13591366

Knowledge Trinity Large Thinking leads

Hunyuan Turbos 20250226: 37.0 (#161), Trinity Large Thinking: 40.9 (#113)

Knowledge benchmarks
BenchmarkHunyuan Turbos 20250226Trinity Large Thinking
LMArena Expert13391360
Vectara Hallucination Rate—6.9%

Multilingual Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 48.9 (#136), Trinity Large Thinking: 46.2 (#160)

Multilingual benchmarks
BenchmarkHunyuan Turbos 20250226Trinity Large Thinking
LMArena Non-English13631325
LMArena Chinese14171373
LMArena French13911374
LMArena German13551356
LMArena Japanese13421311
LMArena Korean13511306
LMArena Russian13681337
LMArena Spanish—1357

Instruction Following Too close to call

Hunyuan Turbos 20250226: 71.0 (#158), Trinity Large Thinking: 70.5 (#162)

Instruction Following benchmarks
BenchmarkHunyuan Turbos 20250226Trinity Large Thinking
LMArena Instruction Following13441334

Long Context Too close to call

Hunyuan Turbos 20250226: 41.6 (#136), Trinity Large Thinking: 41.3 (#144)

Long Context benchmarks
BenchmarkHunyuan Turbos 20250226Trinity Large Thinking
LMArena Longer Query13661355

Writing & Preference Hunyuan Turbos 20250226 leads

Hunyuan Turbos 20250226: 57.4 (#128), Trinity Large Thinking: 53.8 (#158)

Writing & Preference benchmarks
BenchmarkHunyuan Turbos 20250226Trinity Large Thinking
LMArena Text13771340
LMArena Creative Writing13591320
LMArena Multi-Turn13871342

Frequently asked questions

Is Hunyuan Turbos 20250226 better than Trinity Large Thinking?

Hunyuan Turbos 20250226 is the stronger model overall, scoring 41.3 to 38.6 on the Noometry Index.

Is Hunyuan Turbos 20250226 or Trinity Large Thinking better for coding?

Hunyuan Turbos 20250226 scores higher on coding benchmarks: 40.0 versus 34.1 in the Noometry coding category.

How many benchmarks do Hunyuan Turbos 20250226 and Trinity Large Thinking share?

16 benchmarks have published results for both models. Hunyuan Turbos 20250226 has 16 scored results on Noometry and Trinity Large Thinking has 24.

Related comparisons

Go deeper