Model comparison

Grok 4.1 Fast vs Hy3

Hy3 is the stronger model overall, scoring 44.2 to 41.4 on the Noometry Index.

Last verified . 19 shared benchmarks.

Grok 4.1 Fast xAI

41.4

Rank #136 Confirmed

Hy3 Tencent

44.2

Rank #79 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Grok 4.1 Fast scores higher in 1 category and Hy3 in 7 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Grok 4.1 Fast leads 43.4 to 26.1.
  • The biggest single-benchmark swing is NYT Connections (extended): 87.4% for Grok 4.1 Fast and 41.2% for Hy3.
  • Hy3 is cheaper at $0.0825 / $0.33 per million input/output tokens, against $0.20 / $0.50 for Grok 4.1 Fast.
  • Hy3 accepts more context: 262K tokens versus 128K.
  • Hy3 has downloadable open weights; the other is API-only.

Side by side

Grok 4.1 Fast and Hy3 specifications
Grok 4.1 FastHy3
ProviderxAITencent
Noometry Index41.444.2
Released2025-06-272026-07-06
WeightsProprietaryOpen
Context window128K262K
Max output30K128K
Input $ / M tokens$0.20$0.0825
Output $ / M tokens$0.50$0.33
Results tracked3219

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hy3 leads

Grok 4.1 Fast: 34.1 (#245), Hy3: 46.8 (#63)

Coding benchmarks
BenchmarkGrok 4.1 FastHy3
LMArena WebDev12421508
LMArena Coding14111464
ALE-Bench394.93—

Agentic & Tool Use Not comparable

Grok 4.1 Fast: 36.3 (#39), Hy3: —

Agentic & Tool Use benchmarks
BenchmarkGrok 4.1 FastHy3
Berkeley Function Calling Leaderboard69.6%—
τ²-bench Banking13.1%—
LMArena Search1171—
Vending-Bench 21,107—

Reasoning Grok 4.1 Fast leads

Grok 4.1 Fast: 43.4 (#49), Hy3: 26.1 (#136)

Reasoning benchmarks
BenchmarkGrok 4.1 FastHy3
NYT Connections (extended)87.4%41.2%
LMArena Hard Prompts14071447
SimpleBench56%—
DTBench87.7%—
ForecastBench61—

Math Hy3 leads

Grok 4.1 Fast: 31.9 (#221), Hy3: 40.1 (#93)

Math benchmarks
BenchmarkGrok 4.1 FastHy3
LMArena Math14081475
MathArena Final-Answer Competitions60.9%—
ProofBench4%—

Knowledge Hy3 leads

Grok 4.1 Fast: 33.1 (#207), Hy3: 40.8 (#114)

Knowledge benchmarks
BenchmarkGrok 4.1 FastHy3
LMArena Expert13991460
Vectara Hallucination Rate17.8%—

Multimodal Not comparable

Grok 4.1 Fast: 37.0 (#76), Hy3: —

Multimodal benchmarks
BenchmarkGrok 4.1 FastHy3
LMArena Vision1201—

Multilingual Hy3 leads

Grok 4.1 Fast: 51.0 (#114), Hy3: 53.5 (#65)

Multilingual benchmarks
BenchmarkGrok 4.1 FastHy3
LMArena Non-English13911426
LMArena Chinese14411493
LMArena French14151461
LMArena German14041439
LMArena Japanese13491392
LMArena Korean13611395
LMArena Russian13871432
LMArena Spanish14131456

Instruction Following Hy3 leads

Grok 4.1 Fast: 72.7 (#133), Hy3: 75.1 (#70)

Instruction Following benchmarks
BenchmarkGrok 4.1 FastHy3
LMArena Instruction Following13761426

Long Context Hy3 leads

Grok 4.1 Fast: 42.4 (#126), Hy3: 44.1 (#75)

Long Context benchmarks
BenchmarkGrok 4.1 FastHy3
LMArena Longer Query13901442

Writing & Preference Hy3 leads

Grok 4.1 Fast: 57.2 (#131), Hy3: 62.2 (#81)

Writing & Preference benchmarks
BenchmarkGrok 4.1 FastHy3
LMArena Text14081439
LMArena Creative Writing13941402
LMArena Multi-Turn13891436
EQ-Bench Creative Writing1327—

Frequently asked questions

Is Grok 4.1 Fast better than Hy3?

Hy3 is the stronger model overall, scoring 44.2 to 41.4 on the Noometry Index.

Which is cheaper, Grok 4.1 Fast or Hy3?

Hy3 is cheaper. It lists at $0.0825 per million input tokens and $0.33 per million output tokens; Grok 4.1 Fast lists at $0.20 and $0.50.

Is Grok 4.1 Fast or Hy3 better for coding?

Hy3 scores higher on coding benchmarks: 46.8 versus 34.1 in the Noometry coding category.

Which has the bigger context window?

Hy3 does, with 262K tokens against 128K.

How many benchmarks do Grok 4.1 Fast and Hy3 share?

19 benchmarks have published results for both models. Grok 4.1 Fast has 32 scored results on Noometry and Hy3 has 19.

Related comparisons

Go deeper