Model comparison

Hunyuan T1 20250711 vs Olmo 3 32b Think

Hunyuan T1 20250711 is the stronger model overall, scoring 42.5 to 38.7 on the Noometry Index.

Last verified . 12 shared benchmarks.

Hunyuan T1 20250711 Tencent

42.5

Rank #114 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Hunyuan T1 20250711 scores higher in 8 categories and Olmo 3 32b Think in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Hunyuan T1 20250711 leads 59.5 to 49.1.
  • Olmo 3 32b Think has downloadable open weights; the other is API-only.

Side by side

Hunyuan T1 20250711 and Olmo 3 32b Think specifications
Hunyuan T1 20250711Olmo 3 32b Think
ProviderTencentAllen Institute for AI (Ai2)
Noometry Index42.538.7
Released——
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1314

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 40.9 (#129), Olmo 3 32b Think: 38.6 (#172)

Coding benchmarks
BenchmarkHunyuan T1 20250711Olmo 3 32b Think
LMArena Coding13901319

Reasoning Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 28.5 (#103), Olmo 3 32b Think: 25.9 (#140)

Reasoning benchmarks
BenchmarkHunyuan T1 20250711Olmo 3 32b Think
LMArena Hard Prompts13991302

Math Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 38.7 (#130), Olmo 3 32b Think: 36.5 (#165)

Math benchmarks
BenchmarkHunyuan T1 20250711Olmo 3 32b Think
LMArena Math14141316

Knowledge Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 38.8 (#141), Olmo 3 32b Think: 35.0 (#190)

Knowledge benchmarks
BenchmarkHunyuan T1 20250711Olmo 3 32b Think
LMArena Expert13951273

Multilingual Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 51.2 (#112), Olmo 3 32b Think: 41.2 (#210)

Multilingual benchmarks
BenchmarkHunyuan T1 20250711Olmo 3 32b Think
LMArena Non-English13951255
LMArena Chinese14251300
LMArena Russian13851254
LMArena French—1291
LMArena German—1290
LMArena Korean1406—

Instruction Following Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 72.6 (#138), Olmo 3 32b Think: 67.2 (#198)

Instruction Following benchmarks
BenchmarkHunyuan T1 20250711Olmo 3 32b Think
LMArena Instruction Following13741275

Long Context Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 42.2 (#128), Olmo 3 32b Think: 39.4 (#182)

Long Context benchmarks
BenchmarkHunyuan T1 20250711Olmo 3 32b Think
LMArena Longer Query13841296

Writing & Preference Hunyuan T1 20250711 leads

Hunyuan T1 20250711: 59.5 (#109), Olmo 3 32b Think: 49.1 (#193)

Writing & Preference benchmarks
BenchmarkHunyuan T1 20250711Olmo 3 32b Think
LMArena Text14011300
LMArena Creative Writing13921256
LMArena Multi-Turn13931290

Frequently asked questions

Is Hunyuan T1 20250711 better than Olmo 3 32b Think?

Hunyuan T1 20250711 is the stronger model overall, scoring 42.5 to 38.7 on the Noometry Index.

Is Hunyuan T1 20250711 or Olmo 3 32b Think better for coding?

Hunyuan T1 20250711 scores higher on coding benchmarks: 40.9 versus 38.6 in the Noometry coding category.

How many benchmarks do Hunyuan T1 20250711 and Olmo 3 32b Think share?

12 benchmarks have published results for both models. Hunyuan T1 20250711 has 13 scored results on Noometry and Olmo 3 32b Think has 14.

Related comparisons

Go deeper