Model comparison

GLM-4.7-Flash vs Hunyuan Large 2025 02 10

GLM-4.7-Flash and Hunyuan Large 2025 02 10 score almost the same on the Noometry Index (38.8 vs 38.6), so choose on price, context window or the category you care about most.

Last verified . 12 shared benchmarks.

GLM-4.7-Flash Z.ai (Zhipu)

38.8

Rank #180 Confirmed

Hunyuan Large 2025 02 10 Tencent

38.6

Rank #184 Confirmed

Summary

  • They share 12 benchmarks with published results for both. GLM-4.7-Flash scores higher in 6 categories and Hunyuan Large 2025 02 10 in 2 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in multilingual, where GLM-4.7-Flash leads 46.5 to 42.0.
  • GLM-4.7-Flash has downloadable open weights; the other is API-only.

Side by side

GLM-4.7-Flash and Hunyuan Large 2025 02 10 specifications
GLM-4.7-FlashHunyuan Large 2025 02 10
ProviderZ.ai (Zhipu)Tencent
Noometry Index38.838.6
Released2026-01-19—
WeightsOpenProprietary
Context window200K—
Max output131K—
Input $ / M tokens$0.06—
Output $ / M tokens$0.40—
Results tracked2112

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GLM-4.7-Flash leads

GLM-4.7-Flash: 40.6 (#135), Hunyuan Large 2025 02 10: 38.2 (#181)

Coding benchmarks
BenchmarkGLM-4.7-FlashHunyuan Large 2025 02 10
LMArena Coding13831307

Reasoning Hunyuan Large 2025 02 10 leads

GLM-4.7-Flash: 20.9 (#229), Hunyuan Large 2025 02 10: 25.5 (#148)

Reasoning benchmarks
BenchmarkGLM-4.7-FlashHunyuan Large 2025 02 10
LMArena Hard Prompts13561286
Chess Puzzles0%—

Math Too close to call

GLM-4.7-Flash: 36.1 (#173), Hunyuan Large 2025 02 10: 35.8 (#178)

Math benchmarks
BenchmarkGLM-4.7-FlashHunyuan Large 2025 02 10
LMArena Math13551281
OTIS Mock AIME 2024-202558.3%—

Knowledge Too close to call

GLM-4.7-Flash: 35.5 (#184), Hunyuan Large 2025 02 10: 35.1 (#188)

Knowledge benchmarks
BenchmarkGLM-4.7-FlashHunyuan Large 2025 02 10
LMArena Expert13571276
GPQA Diamond60.5%—
Vectara Hallucination Rate9.3%—

Multilingual GLM-4.7-Flash leads

GLM-4.7-Flash: 46.5 (#158), Hunyuan Large 2025 02 10: 42.0 (#200)

Multilingual benchmarks
BenchmarkGLM-4.7-FlashHunyuan Large 2025 02 10
LMArena Non-English13301265
LMArena Chinese14031346
LMArena Russian13321266
LMArena French1332—
LMArena German1337—
LMArena Korean1283—
LMArena Spanish1350—

Instruction Following GLM-4.7-Flash leads

GLM-4.7-Flash: 70.1 (#167), Hunyuan Large 2025 02 10: 67.3 (#197)

Instruction Following benchmarks
BenchmarkGLM-4.7-FlashHunyuan Large 2025 02 10
LMArena Instruction Following13271277

Long Context Too close to call

GLM-4.7-Flash: 40.9 (#148), Hunyuan Large 2025 02 10: 40.8 (#149)

Long Context benchmarks
BenchmarkGLM-4.7-FlashHunyuan Large 2025 02 10
LMArena Longer Query13451341

Writing & Preference Hunyuan Large 2025 02 10 leads

GLM-4.7-Flash: 47.4 (#210), Hunyuan Large 2025 02 10: 48.7 (#197)

Writing & Preference benchmarks
BenchmarkGLM-4.7-FlashHunyuan Large 2025 02 10
LMArena Text13511288
LMArena Creative Writing12971264
LMArena Multi-Turn13421284
EQ-Bench Creative Writing1125—

Frequently asked questions

Is GLM-4.7-Flash better than Hunyuan Large 2025 02 10?

GLM-4.7-Flash and Hunyuan Large 2025 02 10 score almost the same on the Noometry Index (38.8 vs 38.6), so choose on price, context window or the category you care about most.

Is GLM-4.7-Flash or Hunyuan Large 2025 02 10 better for coding?

GLM-4.7-Flash scores higher on coding benchmarks: 40.6 versus 38.2 in the Noometry coding category.

How many benchmarks do GLM-4.7-Flash and Hunyuan Large 2025 02 10 share?

12 benchmarks have published results for both models. GLM-4.7-Flash has 21 scored results on Noometry and Hunyuan Large 2025 02 10 has 12.

Related comparisons

Go deeper