Model comparison
Hy4 preview vs Qwen3.5-9B
Hy4 preview is the stronger model overall, scoring 45.3 to 33.8 on the Noometry Index. Qwen3.5-9B costs 11× less per token, which makes it the better buy when Hy4 preview's lead doesn't matter for your workload.
Last verified . 0 shared benchmarks.
Summary
- The widest gap is in math, where Hy4 preview leads 55.7 to 34.8.
- Qwen3.5-9B is cheaper at $0.10 / $0.15 per million input/output tokens, against $0.83 / $2.50 for Hy4 preview.
- Hy4 preview accepts more context: 1.05M tokens versus 262K.
Side by side
| Hy4 preview | Qwen3.5-9B | |
|---|---|---|
| Provider | Tencent | Alibaba (Qwen) |
| Noometry Index | 45.3 | 33.8 |
| Released | 2026-08-28 | 2026-02-23 |
| Weights | Open | Open |
| Context window | 1.05M | 262K |
| Max output | 64K | 66K |
| Input $ / M tokens | $0.83 | $0.10 |
| Output $ / M tokens | $2.50 | $0.15 |
| Results tracked | 3 | 10 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Hy4 preview leads
Hy4 preview: 51.6 (#38), Qwen3.5-9B: 35.9 (#217)
| Benchmark | Hy4 preview | Qwen3.5-9B |
|---|---|---|
| LMArena WebDev | 1632 | — |
| SciCode | — | 27.5% |
Agentic & Tool Use Not comparable
Hy4 preview: —, Qwen3.5-9B: 14.5 (#151)
| Benchmark | Hy4 preview | Qwen3.5-9B |
|---|---|---|
| Terminal-Bench | — | 9.2% |
Reasoning Hy4 preview leads
Hy4 preview: 31.9 (#79), Qwen3.5-9B: 23.1 (#182)
| Benchmark | Hy4 preview | Qwen3.5-9B |
|---|---|---|
| NYT Connections (extended) | 68.2% | — |
| CritPt | — | 0.3% |
| Chess Puzzles | — | 12% |
| DTBench | — | 71.2% |
| LMCA | — | 24.5% |
| Epoch Capabilities Index | — | 139.46 |
Math Hy4 preview leads
Hy4 preview: 55.7 (#42), Qwen3.5-9B: 34.8 (#192)
| Benchmark | Hy4 preview | Qwen3.5-9B |
|---|---|---|
| MathArena Final-Answer Competitions | — | 48.5% |
| OTIS Mock AIME 2024-2025 | — | 61.7% |
| ProofBench | 75% | — |
Knowledge Not comparable
Hy4 preview: —, Qwen3.5-9B: 46.0 (#84)
| Benchmark | Hy4 preview | Qwen3.5-9B |
|---|---|---|
| GPQA Diamond | — | 79% |
Frequently asked questions
Is Hy4 preview better than Qwen3.5-9B?
Hy4 preview is the stronger model overall, scoring 45.3 to 33.8 on the Noometry Index. Qwen3.5-9B costs 11× less per token, which makes it the better buy when Hy4 preview's lead doesn't matter for your workload.
Which is cheaper, Hy4 preview or Qwen3.5-9B?
Qwen3.5-9B is cheaper. It lists at $0.10 per million input tokens and $0.15 per million output tokens; Hy4 preview lists at $0.83 and $2.50.
Is Hy4 preview or Qwen3.5-9B better for coding?
Hy4 preview scores higher on coding benchmarks: 51.6 versus 35.9 in the Noometry coding category.
Which has the bigger context window?
Hy4 preview does, with 1.05M tokens against 262K.
How many benchmarks do Hy4 preview and Qwen3.5-9B share?
0 benchmarks have published results for both models. Hy4 preview has 3 scored results on Noometry and Qwen3.5-9B has 10.