Model comparison

Hy4 preview vs Mistral 7B

Hy4 preview is the stronger model overall, scoring 45.3 to 23.0 on the Noometry Index. Mistral 7B costs 4.5× less per token, which makes it the better buy when Hy4 preview's lead doesn't matter for your workload.

Last verified . 0 shared benchmarks.

Hy4 preview Tencent

45.3

Rank #73 Reported

Mistral 7B Mistral AI

23.0

Rank #351 Confirmed

Summary

  • The widest gap is in math, where Hy4 preview leads 55.7 to 8.1.
  • Mistral 7B is cheaper at $0.25 / $0.25 per million input/output tokens, against $0.75 / $2.25 for Hy4 preview.
  • Hy4 preview accepts more context: 1.05M tokens versus 8K.

Side by side

Hy4 preview and Mistral 7B specifications
Hy4 previewMistral 7B
ProviderTencentMistral AI
Noometry Index45.323.0
Released2026-08-282023-09-27
WeightsOpenOpen
Context window1.05M8K
Max output64K8K
Input $ / M tokens$0.75$0.25
Output $ / M tokens$2.25$0.25
Results tracked337

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hy4 preview leads

Hy4 preview: 51.6 (#38), Mistral 7B: 26.4 (#326)

Coding benchmarks
BenchmarkHy4 previewMistral 7B
LMArena WebDev1632—
BigCodeBench Instruct—19.5%
LMArena Coding—1082
BigCodeBench Complete—27.3%
HumanEval+—36%
MBPP+—42.1%

Reasoning Hy4 preview leads

Hy4 preview: 31.9 (#79), Mistral 7B: 13.1 (#336)

Reasoning benchmarks
BenchmarkHy4 previewMistral 7B
NYT Connections (extended)68.2%—
Chess Puzzles—0%
LMArena Hard Prompts—1067
DTBench—42.5%
Adversarial NLI—47.1%
BIG-Bench Hard—56.1%
Epoch Capabilities Index—112.21
HellaSwag—81%
PIQA—83%
WinoGrande—75.3%

Math Hy4 preview leads

Hy4 preview: 55.7 (#42), Mistral 7B: 8.1 (#325)

Math benchmarks
BenchmarkHy4 previewMistral 7B
OTIS Mock AIME 2024-2025—0.3%
ProofBench75%—
LMArena Math—1085
MATH Level 5—3.7%
GSM8K—54.4%

Knowledge Not comparable

Hy4 preview: —, Mistral 7B: 7.4 (#311)

Knowledge benchmarks
BenchmarkHy4 previewMistral 7B
GPQA Diamond—15.2%
LMArena Expert—1036
ARC (AI2) Challenge—78.6%
BoolQ—87.4%
MMLU—62.5%
OpenBookQA—79.8%
TriviaQA—75.2%

Multilingual Not comparable

Hy4 preview: —, Mistral 7B: 25.8 (#283)

Multilingual benchmarks
BenchmarkHy4 previewMistral 7B
LMArena Non-English—1012
LMArena Chinese—1009
LMArena French—1037
LMArena German—987
LMArena Japanese—878
LMArena Russian—1018
LMArena Spanish—1026

Instruction Following Not comparable

Hy4 preview: —, Mistral 7B: 54.2 (#280)

Instruction Following benchmarks
BenchmarkHy4 previewMistral 7B
LMArena Instruction Following—1060

Long Context Not comparable

Hy4 preview: —, Mistral 7B: 32.2 (#271)

Long Context benchmarks
BenchmarkHy4 previewMistral 7B
LMArena Longer Query—1060

Writing & Preference Not comparable

Hy4 preview: —, Mistral 7B: 30.7 (#286)

Writing & Preference benchmarks
BenchmarkHy4 previewMistral 7B
LMArena Text—1090
LMArena Creative Writing—1068
LMArena Multi-Turn—1062

Frequently asked questions

Is Hy4 preview better than Mistral 7B?

Hy4 preview is the stronger model overall, scoring 45.3 to 23.0 on the Noometry Index. Mistral 7B costs 4.5× less per token, which makes it the better buy when Hy4 preview's lead doesn't matter for your workload.

Which is cheaper, Hy4 preview or Mistral 7B?

Mistral 7B is cheaper. It lists at $0.25 per million input tokens and $0.25 per million output tokens; Hy4 preview lists at $0.75 and $2.25.

Is Hy4 preview or Mistral 7B better for coding?

Hy4 preview scores higher on coding benchmarks: 51.6 versus 26.4 in the Noometry coding category.

Which has the bigger context window?

Hy4 preview does, with 1.05M tokens against 8K.

How many benchmarks do Hy4 preview and Mistral 7B share?

0 benchmarks have published results for both models. Hy4 preview has 3 scored results on Noometry and Mistral 7B has 37.

Related comparisons

Go deeper