Model comparison

Grok 4.7 vs Hunyuan Vision 1.5

Grok 4.7 is the stronger model overall, scoring 53.1 to 43.1 on the Noometry Index.

Last verified . 9 shared benchmarks.

Grok 4.7 xAI

53.1

Rank #37 Confirmed

Hunyuan Vision 1.5 Tencent

43.1

Rank #98 Confirmed

Summary

  • They share 9 benchmarks with published results for both. Grok 4.7 scores higher in 6 categories and Hunyuan Vision 1.5 in 1 category; 3 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Grok 4.7 leads 49.1 to 28.9.

Side by side

Grok 4.7 and Hunyuan Vision 1.5 specifications
Grok 4.7Hunyuan Vision 1.5
ProviderxAITencent
Noometry Index53.143.1
Released2026-09-21—
WeightsProprietaryProprietary
Context window500K—
Max output500K—
Input $ / M tokens$2—
Output $ / M tokens$6—
Results tracked399

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Grok 4.7 leads

Grok 4.7: 58.0 (#18), Hunyuan Vision 1.5: 41.8 (#119)

Coding benchmarks
BenchmarkGrok 4.7Hunyuan Vision 1.5
LMArena Coding14271418
FrontierCode47.6%—
CursorBench46.3%—
LMArena WebDev1639—
FrontierSWE29.5%—
SciCode57.8%—

Agentic & Tool Use Not comparable

Grok 4.7: 36.7 (#37), Hunyuan Vision 1.5: —

Agentic & Tool Use benchmarks
BenchmarkGrok 4.7Hunyuan Vision 1.5
APEX-Agents54.6%—
GDP.pdf22.8%—
Vending-Bench 210,537—

Reasoning Grok 4.7 leads

Grok 4.7: 49.1 (#40), Hunyuan Vision 1.5: 28.9 (#97)

Reasoning benchmarks
BenchmarkGrok 4.7Hunyuan Vision 1.5
LMArena Hard Prompts14131416
NYT Connections (extended)76.8%—
CritPt18%—
Chess Puzzles38%—
Mystery Game Puzzles29%—
DTBench96%—
LMCA49.4%—
Epoch Capabilities Index153.53—

Math Not comparable

Grok 4.7: 57.8 (#39), Hunyuan Vision 1.5: —

Math benchmarks
BenchmarkGrok 4.7Hunyuan Vision 1.5
FrontierMath (Tiers 1-3)53%—
FrontierMath Tier 417.1%—
OTIS Mock AIME 2024-202598.1%—
ProofBench34%—
LMArena Math1407—

Knowledge Not comparable

Grok 4.7: 62.8 (#22), Hunyuan Vision 1.5: —

Knowledge benchmarks
BenchmarkGrok 4.7Hunyuan Vision 1.5
GPQA Diamond92.7%—
SimpleQA Verified56%—
LMArena Expert1422—

Multimodal Too close to call

Grok 4.7: 35.5 (#87), Hunyuan Vision 1.5: 35.7 (#84)

Multimodal benchmarks
BenchmarkGrok 4.7Hunyuan Vision 1.5
LMArena Vision12281180
Blueprint-Bench 232.5%—
Furniture Assembly20.8%—

Multilingual Too close to call

Grok 4.7: 50.8 (#116), Hunyuan Vision 1.5: 50.3 (#125)

Multilingual benchmarks
BenchmarkGrok 4.7Hunyuan Vision 1.5
LMArena Non-English13891382
LMArena Chinese1455—
LMArena French1455—
LMArena Russian1397—
LMArena Spanish1400—

Instruction Following Too close to call

Grok 4.7: 74.1 (#105), Hunyuan Vision 1.5: 73.7 (#117)

Instruction Following benchmarks
BenchmarkGrok 4.7Hunyuan Vision 1.5
LMArena Instruction Following14041397

Long Context Too close to call

Grok 4.7: 43.1 (#104), Hunyuan Vision 1.5: 42.9 (#115)

Long Context benchmarks
BenchmarkGrok 4.7Hunyuan Vision 1.5
LMArena Longer Query14131404

Writing & Preference Grok 4.7 leads

Grok 4.7: 70.0 (#24), Hunyuan Vision 1.5: 60.1 (#100)

Writing & Preference benchmarks
BenchmarkGrok 4.7Hunyuan Vision 1.5
LMArena Text13991405
LMArena Creative Writing13911388
LMArena Multi-Turn13931422
EQ-Bench Creative Writing2007—

Frequently asked questions

Is Grok 4.7 better than Hunyuan Vision 1.5?

Grok 4.7 is the stronger model overall, scoring 53.1 to 43.1 on the Noometry Index.

Is Grok 4.7 or Hunyuan Vision 1.5 better for coding?

Grok 4.7 scores higher on coding benchmarks: 58.0 versus 41.8 in the Noometry coding category.

How many benchmarks do Grok 4.7 and Hunyuan Vision 1.5 share?

9 benchmarks have published results for both models. Grok 4.7 has 39 scored results on Noometry and Hunyuan Vision 1.5 has 9.

Related comparisons

Go deeper