Model comparison

Claude 2.1 vs Hy4 preview

Hy4 preview is the stronger model overall, scoring 45.3 to 25.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Claude 2.1 Anthropic

25.2

Rank #345 Reported

Hy4 preview Tencent

45.3

Rank #73 Reported

Summary

  • The widest gap is in math, where Hy4 preview leads 55.7 to 10.2.
  • Hy4 preview has downloadable open weights; the other is API-only.

Side by side

Claude 2.1 and Hy4 preview specifications
Claude 2.1Hy4 preview
ProviderAnthropicTencent
Noometry Index25.245.3
Released2023-11-212026-08-28
WeightsProprietaryOpen
Context window—1.05M
Max output—64K
Input $ / M tokens—$0.75
Output $ / M tokens—$2.25
Results tracked73

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hy4 preview leads

Claude 2.1: 26.2 (#327), Hy4 preview: 51.6 (#38)

Coding benchmarks
BenchmarkClaude 2.1Hy4 preview
LMArena WebDev—1632
WeirdML7.1%—

Reasoning Hy4 preview leads

Claude 2.1: 21.4 (#221), Hy4 preview: 31.9 (#79)

Reasoning benchmarks
BenchmarkClaude 2.1Hy4 preview
NYT Connections (extended)—68.2%
DTBench51%—
Epoch Capabilities Index119.27—
ForecastBench54.2—

Math Hy4 preview leads

Claude 2.1: 10.2 (#315), Hy4 preview: 55.7 (#42)

Math benchmarks
BenchmarkClaude 2.1Hy4 preview
OTIS Mock AIME 2024-20251.9%—
ProofBench—75%

Knowledge Not comparable

Claude 2.1: 15.4 (#292), Hy4 preview: —

Knowledge benchmarks
BenchmarkClaude 2.1Hy4 preview
GPQA Diamond33%—
MMLU73.5%—

Frequently asked questions

Is Claude 2.1 better than Hy4 preview?

Hy4 preview is the stronger model overall, scoring 45.3 to 25.2 on the Noometry Index.

Is Claude 2.1 or Hy4 preview better for coding?

Hy4 preview scores higher on coding benchmarks: 51.6 versus 26.2 in the Noometry coding category.

How many benchmarks do Claude 2.1 and Hy4 preview share?

0 benchmarks have published results for both models. Claude 2.1 has 7 scored results on Noometry and Hy4 preview has 3.

Related comparisons

Go deeper