Model comparison

Claude 3 Sonnet vs Hy3

Hy3 is the stronger model overall, scoring 44.2 to 29.0 on the Noometry Index.

Last verified . 17 shared benchmarks.

Claude 3 Sonnet Anthropic

29.0

Rank #319 Confirmed

Hy3 Tencent

44.2

Rank #79 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Claude 3 Sonnet scores higher in 0 categories and Hy3 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in math, where Hy3 leads 40.1 to 10.7.
  • Hy3 has downloadable open weights; the other is API-only.

Side by side

Claude 3 Sonnet and Hy3 specifications
Claude 3 SonnetHy3
ProviderAnthropicTencent
Noometry Index29.044.2
Released2024-02-292026-07-06
WeightsProprietaryOpen
Context window—262K
Max output—128K
Input $ / M tokens—$0.0825
Output $ / M tokens—$0.33
Results tracked3019

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hy3 leads

Claude 3 Sonnet: 29.6 (#302), Hy3: 46.8 (#63)

Coding benchmarks
BenchmarkClaude 3 SonnetHy3
LMArena Coding12231464
LMArena WebDev—1508
WeirdML10.2%—
BigCodeBench Instruct42.7%—
BigCodeBench Complete53.8%—
HumanEval+64%—
MBPP+69.3%—

Reasoning Hy3 leads

Claude 3 Sonnet: 20.5 (#237), Hy3: 26.1 (#136)

Reasoning benchmarks
BenchmarkClaude 3 SonnetHy3
LMArena Hard Prompts11971447
NYT Connections (extended)—41.2%
DTBench53.6%—
Epoch Capabilities Index120.7—
WinoGrande75.1%—

Math Hy3 leads

Claude 3 Sonnet: 10.7 (#310), Hy3: 40.1 (#93)

Math benchmarks
BenchmarkClaude 3 SonnetHy3
LMArena Math12131475
OTIS Mock AIME 2024-20252.5%—
MATH Level 518.2%—

Knowledge Hy3 leads

Claude 3 Sonnet: 21.1 (#276), Hy3: 40.8 (#114)

Knowledge benchmarks
BenchmarkClaude 3 SonnetHy3
LMArena Expert11731460
GPQA Diamond40.6%—
MMLU75.9%—

Multimodal Not comparable

Claude 3 Sonnet: 25.2 (#125), Hy3: —

Multimodal benchmarks
BenchmarkClaude 3 SonnetHy3
LMArena Vision984—

Multilingual Hy3 leads

Claude 3 Sonnet: 37.8 (#234), Hy3: 53.5 (#65)

Multilingual benchmarks
BenchmarkClaude 3 SonnetHy3
LMArena Non-English12051426
LMArena Chinese11891493
LMArena French12291461
LMArena German12041439
LMArena Japanese11311392
LMArena Korean11281395
LMArena Russian12271432
LMArena Spanish12041456

Instruction Following Hy3 leads

Claude 3 Sonnet: 62.8 (#235), Hy3: 75.1 (#70)

Instruction Following benchmarks
BenchmarkClaude 3 SonnetHy3
LMArena Instruction Following11991426

Long Context Hy3 leads

Claude 3 Sonnet: 36.7 (#228), Hy3: 44.1 (#75)

Long Context benchmarks
BenchmarkClaude 3 SonnetHy3
LMArena Longer Query12111442

Writing & Preference Hy3 leads

Claude 3 Sonnet: 42.1 (#238), Hy3: 62.2 (#81)

Writing & Preference benchmarks
BenchmarkClaude 3 SonnetHy3
LMArena Text12181439
LMArena Creative Writing11861402
LMArena Multi-Turn12271436

Frequently asked questions

Is Claude 3 Sonnet better than Hy3?

Hy3 is the stronger model overall, scoring 44.2 to 29.0 on the Noometry Index.

Is Claude 3 Sonnet or Hy3 better for coding?

Hy3 scores higher on coding benchmarks: 46.8 versus 29.6 in the Noometry coding category.

How many benchmarks do Claude 3 Sonnet and Hy3 share?

17 benchmarks have published results for both models. Claude 3 Sonnet has 30 scored results on Noometry and Hy3 has 19.

Related comparisons

Go deeper