Model comparison

Hy3 vs Mistral Medium

Hy3 is the stronger model overall, scoring 44.2 to 36.3 on the Noometry Index.

Last verified . 17 shared benchmarks.

Hy3 Tencent

44.2

Rank #79 Confirmed

Mistral Medium Mistral AI

36.3

Rank #218 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Hy3 scores higher in 8 categories and Mistral Medium in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Hy3 leads 40.8 to 25.0.
  • Hy3 is cheaper at $0.0825 / $0.33 per million input/output tokens, against $1.50 / $7.50 for Mistral Medium.

Side by side

Hy3 and Mistral Medium specifications
Hy3Mistral Medium
ProviderTencentMistral AI
Noometry Index44.236.3
Released2026-07-062023-12-11
WeightsOpenOpen
Context window262K262K
Max output128K262K
Input $ / M tokens$0.0825$1.50
Output $ / M tokens$0.33$7.50
Results tracked1936

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hy3 leads

Hy3: 46.8 (#63), Mistral Medium: 34.2 (#243)

Coding benchmarks
BenchmarkHy3Mistral Medium
LMArena Coding14641434
FrontierCode—8%
LMArena WebDev1508—
SciCode—40.2%
WeirdML—43.7%
ALE-Bench—763.98

Agentic & Tool Use Not comparable

Hy3: —, Mistral Medium: 28.3 (#90)

Agentic & Tool Use benchmarks
BenchmarkHy3Mistral Medium
Berkeley Function Calling Leaderboard—37.7%

Reasoning Hy3 leads

Hy3: 26.1 (#136), Mistral Medium: 24.0 (#167)

Reasoning benchmarks
BenchmarkHy3Mistral Medium
LMArena Hard Prompts14471426
Kagi LLM Benchmark—50%
NYT Connections (extended)41.2%—
CritPt—0%
DTBench—75.5%
LMCA—26.1%
Surface Evolver Bench—26.9%

Math Hy3 leads

Hy3: 40.1 (#93), Mistral Medium: 28.1 (#245)

Math benchmarks
BenchmarkHy3Mistral Medium
LMArena Math14751408
OTIS Mock AIME 2024-2025—32.2%
ProofBench—9%
MATH Level 5—81.6%
FrontierMath (Feb 2025 set)—0.3%

Knowledge Hy3 leads

Hy3: 40.8 (#114), Mistral Medium: 25.0 (#265)

Knowledge benchmarks
BenchmarkHy3Mistral Medium
LMArena Expert14601408
GPQA Diamond—59.5%
Humanity's Last Exam—4.5%
Vectara Hallucination Rate—22.7%

Multimodal Not comparable

Hy3: —, Mistral Medium: 35.3 (#88)

Multimodal benchmarks
BenchmarkHy3Mistral Medium
LMArena Vision—1172

Multilingual Hy3 leads

Hy3: 53.5 (#65), Mistral Medium: 52.1 (#91)

Multilingual benchmarks
BenchmarkHy3Mistral Medium
LMArena Non-English14261408
LMArena Chinese14931447
LMArena French14611459
LMArena German14391432
LMArena Japanese13921378
LMArena Korean13951380
LMArena Russian14321411
LMArena Spanish14561433

Instruction Following Hy3 leads

Hy3: 75.1 (#70), Mistral Medium: 73.7 (#116)

Instruction Following benchmarks
BenchmarkHy3Mistral Medium
LMArena Instruction Following14261398

Long Context Hy3 leads

Hy3: 44.1 (#75), Mistral Medium: 42.9 (#114)

Long Context benchmarks
BenchmarkHy3Mistral Medium
LMArena Longer Query14421406

Writing & Preference Hy3 leads

Hy3: 62.2 (#81), Mistral Medium: 60.0 (#103)

Writing & Preference benchmarks
BenchmarkHy3Mistral Medium
LMArena Text14391424
LMArena Creative Writing14021391
LMArena Multi-Turn14361418
Short-Story Creative Writing—77.3%

Frequently asked questions

Is Hy3 better than Mistral Medium?

Hy3 is the stronger model overall, scoring 44.2 to 36.3 on the Noometry Index.

Which is cheaper, Hy3 or Mistral Medium?

Hy3 is cheaper. It lists at $0.0825 per million input tokens and $0.33 per million output tokens; Mistral Medium lists at $1.50 and $7.50.

Is Hy3 or Mistral Medium better for coding?

Hy3 scores higher on coding benchmarks: 46.8 versus 34.2 in the Noometry coding category.

Which has the bigger context window?

Both accept 262K tokens.

How many benchmarks do Hy3 and Mistral Medium share?

17 benchmarks have published results for both models. Hy3 has 19 scored results on Noometry and Mistral Medium has 36.

Related comparisons

Go deeper