Model comparison

Hy3 vs Qwen3-Next 80B-A3B Instruct

Hy3 is the stronger model overall, scoring 44.2 to 43.0 on the Noometry Index.

Last verified . 17 shared benchmarks.

Hy3 Tencent

44.2

Rank #79 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Hy3 scores higher in 6 categories and Qwen3-Next 80B-A3B Instruct in 2 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in long context, where Hy3 leads 44.1 to 37.0.
  • Hy3 is cheaper at $0.13 / $0.53 per million input/output tokens, against $0.50 / $2 for Qwen3-Next 80B-A3B Instruct.
  • Hy3 accepts more context: 262K tokens versus 131K.

Side by side

Hy3 and Qwen3-Next 80B-A3B Instruct specifications
Hy3Qwen3-Next 80B-A3B Instruct
ProviderTencentAlibaba (Qwen)
Noometry Index44.243.0
Released2026-07-062025-09
WeightsOpenOpen
Context window262K131K
Max output128K33K
Input $ / M tokens$0.13$0.50
Output $ / M tokens$0.53$2
Results tracked1925

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hy3 leads

Hy3: 46.8 (#63), Qwen3-Next 80B-A3B Instruct: 42.5 (#98)

Coding benchmarks
BenchmarkHy3Qwen3-Next 80B-A3B Instruct
LMArena Coding14641440
LMArena WebDev1508—

Reasoning Qwen3-Next 80B-A3B Instruct leads

Hy3: 26.1 (#136), Qwen3-Next 80B-A3B Instruct: 31.1 (#81)

Reasoning benchmarks
BenchmarkHy3Qwen3-Next 80B-A3B Instruct
LMArena Hard Prompts14471428
Kagi LLM Benchmark—66.7%
NYT Connections (extended)41.2%—

Math Hy3 leads

Hy3: 40.1 (#93), Qwen3-Next 80B-A3B Instruct: 38.8 (#126)

Math benchmarks
BenchmarkHy3Qwen3-Next 80B-A3B Instruct
LMArena Math14751440
Omni-MATH—46.7%

Knowledge Too close to call

Hy3: 40.8 (#114), Qwen3-Next 80B-A3B Instruct: 41.8 (#106)

Knowledge benchmarks
BenchmarkHy3Qwen3-Next 80B-A3B Instruct
LMArena Expert14601417
MMLU-Pro—78.6%
Vectara Hallucination Rate—9.3%
GPQA (HELM)—63%

Multilingual Hy3 leads

Hy3: 53.5 (#65), Qwen3-Next 80B-A3B Instruct: 52.1 (#93)

Multilingual benchmarks
BenchmarkHy3Qwen3-Next 80B-A3B Instruct
LMArena Non-English14261407
LMArena Chinese14931460
LMArena French14611413
LMArena German14391417
LMArena Japanese13921395
LMArena Korean13951364
LMArena Russian14321404
LMArena Spanish14561435

Instruction Following Hy3 leads

Hy3: 75.1 (#70), Qwen3-Next 80B-A3B Instruct: 70.8 (#159)

Instruction Following benchmarks
BenchmarkHy3Qwen3-Next 80B-A3B Instruct
LMArena Instruction Following14261389
IFEval—81%

Long Context Hy3 leads

Hy3: 44.1 (#75), Qwen3-Next 80B-A3B Instruct: 37.0 (#223)

Long Context benchmarks
BenchmarkHy3Qwen3-Next 80B-A3B Instruct
LMArena Longer Query14421403
Fiction.LiveBench—55.6%

Writing & Preference Hy3 leads

Hy3: 62.2 (#81), Qwen3-Next 80B-A3B Instruct: 58.0 (#121)

Writing & Preference benchmarks
BenchmarkHy3Qwen3-Next 80B-A3B Instruct
LMArena Text14391417
LMArena Creative Writing14021334
LMArena Multi-Turn14361416
WildBench—80.7%

Frequently asked questions

Is Hy3 better than Qwen3-Next 80B-A3B Instruct?

Hy3 is the stronger model overall, scoring 44.2 to 43.0 on the Noometry Index.

Which is cheaper, Hy3 or Qwen3-Next 80B-A3B Instruct?

Hy3 is cheaper. It lists at $0.13 per million input tokens and $0.53 per million output tokens; Qwen3-Next 80B-A3B Instruct lists at $0.50 and $2.

Is Hy3 or Qwen3-Next 80B-A3B Instruct better for coding?

Hy3 scores higher on coding benchmarks: 46.8 versus 42.5 in the Noometry coding category.

Which has the bigger context window?

Hy3 does, with 262K tokens against 131K.

How many benchmarks do Hy3 and Qwen3-Next 80B-A3B Instruct share?

17 benchmarks have published results for both models. Hy3 has 19 scored results on Noometry and Qwen3-Next 80B-A3B Instruct has 25.

Related comparisons

Go deeper