Model comparison

Hunyuan Standard 2025 02 10 vs Olmo 3.1 32b Instruct

Olmo 3.1 32b Instruct is the stronger model overall, scoring 39.4 to 37.9 on the Noometry Index.

Last verified . 12 shared benchmarks.

Summary

  • They share 12 benchmarks with published results for both. Hunyuan Standard 2025 02 10 scores higher in 0 categories and Olmo 3.1 32b Instruct in 8 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in instruction following, where Olmo 3.1 32b Instruct leads 68.6 to 65.5.
  • Olmo 3.1 32b Instruct has downloadable open weights; the other is API-only.

Side by side

Hunyuan Standard 2025 02 10 and Olmo 3.1 32b Instruct specifications
Hunyuan Standard 2025 02 10Olmo 3.1 32b Instruct
ProviderTencentAllen Institute for AI (Ai2)
Noometry Index37.939.4
Released——
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1216

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Olmo 3.1 32b Instruct leads

Hunyuan Standard 2025 02 10: 37.1 (#197), Olmo 3.1 32b Instruct: 39.5 (#157)

Coding benchmarks
BenchmarkHunyuan Standard 2025 02 10Olmo 3.1 32b Instruct
LMArena Coding12701347

Reasoning Olmo 3.1 32b Instruct leads

Hunyuan Standard 2025 02 10: 25.0 (#154), Olmo 3.1 32b Instruct: 26.4 (#132)

Reasoning benchmarks
BenchmarkHunyuan Standard 2025 02 10Olmo 3.1 32b Instruct
LMArena Hard Prompts12641322

Math Too close to call

Hunyuan Standard 2025 02 10: 35.6 (#179), Olmo 3.1 32b Instruct: 36.3 (#167)

Math benchmarks
BenchmarkHunyuan Standard 2025 02 10Olmo 3.1 32b Instruct
LMArena Math12741305

Knowledge Olmo 3.1 32b Instruct leads

Hunyuan Standard 2025 02 10: 34.2 (#197), Olmo 3.1 32b Instruct: 36.1 (#175)

Knowledge benchmarks
BenchmarkHunyuan Standard 2025 02 10Olmo 3.1 32b Instruct
LMArena Expert12481308

Multilingual Too close to call

Hunyuan Standard 2025 02 10: 41.7 (#203), Olmo 3.1 32b Instruct: 42.6 (#191)

Multilingual benchmarks
BenchmarkHunyuan Standard 2025 02 10Olmo 3.1 32b Instruct
LMArena Non-English12621275
LMArena Chinese13191304
LMArena Russian12581268
LMArena French—1328
LMArena German—1282
LMArena Korean—1206
LMArena Spanish—1336

Instruction Following Olmo 3.1 32b Instruct leads

Hunyuan Standard 2025 02 10: 65.5 (#219), Olmo 3.1 32b Instruct: 68.6 (#187)

Instruction Following benchmarks
BenchmarkHunyuan Standard 2025 02 10Olmo 3.1 32b Instruct
LMArena Instruction Following12451299

Long Context Too close to call

Hunyuan Standard 2025 02 10: 39.5 (#173), Olmo 3.1 32b Instruct: 39.9 (#166)

Long Context benchmarks
BenchmarkHunyuan Standard 2025 02 10Olmo 3.1 32b Instruct
LMArena Longer Query13011312

Writing & Preference Olmo 3.1 32b Instruct leads

Hunyuan Standard 2025 02 10: 47.2 (#214), Olmo 3.1 32b Instruct: 50.2 (#185)

Writing & Preference benchmarks
BenchmarkHunyuan Standard 2025 02 10Olmo 3.1 32b Instruct
LMArena Text12741311
LMArena Creative Writing12421264
LMArena Multi-Turn12751309

Frequently asked questions

Is Hunyuan Standard 2025 02 10 better than Olmo 3.1 32b Instruct?

Olmo 3.1 32b Instruct is the stronger model overall, scoring 39.4 to 37.9 on the Noometry Index.

Is Hunyuan Standard 2025 02 10 or Olmo 3.1 32b Instruct better for coding?

Olmo 3.1 32b Instruct scores higher on coding benchmarks: 39.5 versus 37.1 in the Noometry coding category.

How many benchmarks do Hunyuan Standard 2025 02 10 and Olmo 3.1 32b Instruct share?

12 benchmarks have published results for both models. Hunyuan Standard 2025 02 10 has 12 scored results on Noometry and Olmo 3.1 32b Instruct has 16.

Related comparisons

Go deeper