Model comparison

GPT-5 Mini vs Hunyuan T1 20250711

GPT-5 Mini and Hunyuan T1 20250711 score almost the same on the Noometry Index (41.8 vs 42.5), so choose on price, context window or the category you care about most.

Last verified . 13 shared benchmarks.

GPT-5 Mini OpenAI

41.8

Rank #128 Confirmed

Hunyuan T1 20250711 Tencent

42.5

Rank #114 Confirmed

Summary

  • They share 13 benchmarks with published results for both. GPT-5 Mini scores higher in 3 categories and Hunyuan T1 20250711 in 5 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in math, where GPT-5 Mini leads 46.7 to 38.7.

Side by side

GPT-5 Mini and Hunyuan T1 20250711 specifications
GPT-5 MiniHunyuan T1 20250711
ProviderOpenAITencent
Noometry Index41.842.5
Released2025-08-07—
WeightsProprietaryProprietary
Context window400K—
Max output128K—
Input $ / M tokens$0.25—
Output $ / M tokens$2—
Results tracked6013

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

GPT-5 Mini: 40.1 (#146), Hunyuan T1 20250711: 40.9 (#129)

Coding benchmarks
BenchmarkGPT-5 MiniHunyuan T1 20250711
LMArena Coding14061390
SWE-bench Verified64.7%—
SWE-bench Verified (bash only)59.8%—
SWE-bench Multilingual39.7%—
SciCode39.2%—
WeirdML52.7%—
ALE-Bench799.77—
AlgoTune1.38—

Agentic & Tool Use Not comparable

GPT-5 Mini: 31.1 (#70), Hunyuan T1 20250711: —

Agentic & Tool Use benchmarks
BenchmarkGPT-5 MiniHunyuan T1 20250711
Terminal-Bench34.8%—
Berkeley Function Calling Leaderboard55.5%—
Vending-Bench 2-31.18—

Reasoning Hunyuan T1 20250711 leads

GPT-5 Mini: 23.9 (#168), Hunyuan T1 20250711: 28.5 (#103)

Reasoning benchmarks
BenchmarkGPT-5 MiniHunyuan T1 20250711
LMArena Hard Prompts13801399
ARC-AGI-24.4%—
Kagi LLM Benchmark70.3%—
ARC-AGI-154.3%—
CritPt0%—
Chess Puzzles30%—
EnigmaEval8.2%—
Mystery Game Puzzles10%—
DTBench80.5%—
LMCA34.2%—
Epoch Capabilities Index145.52—
ForecastBench61—

Math GPT-5 Mini leads

GPT-5 Mini: 46.7 (#69), Hunyuan T1 20250711: 38.7 (#130)

Math benchmarks
BenchmarkGPT-5 MiniHunyuan T1 20250711
LMArena Math13781414
FrontierMath (Tiers 1-3)46.7%—
FrontierMath Tier 412.2%—
OTIS Mock AIME 2024-202586.7%—
ProofBench9%—
Omni-MATH72.2%—
MATH Level 597.8%—
FrontierMath (Feb 2025 set)27.2%—
FrontierMath Tier 4 (v1)6.3%—

Knowledge GPT-5 Mini leads

GPT-5 Mini: 45.6 (#86), Hunyuan T1 20250711: 38.8 (#141)

Knowledge benchmarks
BenchmarkGPT-5 MiniHunyuan T1 20250711
LMArena Expert13791395
GPQA Diamond75%—
Humanity's Last Exam19.4%—
SimpleQA Verified21.6%—
MMLU-Pro83.5%—
Confabulations13.3%—
Vectara Hallucination Rate12.9%—
GPQA (HELM)75.6%—

Multimodal Not comparable

GPT-5 Mini: 35.6 (#85), Hunyuan T1 20250711: —

Multimodal benchmarks
BenchmarkGPT-5 MiniHunyuan T1 20250711
LMArena Vision1202—
VPCT40.2%—

Multilingual Hunyuan T1 20250711 leads

GPT-5 Mini: 48.9 (#137), Hunyuan T1 20250711: 51.2 (#112)

Multilingual benchmarks
BenchmarkGPT-5 MiniHunyuan T1 20250711
LMArena Non-English13631395
LMArena Chinese13851425
LMArena Korean13081406
LMArena Russian13621385
LMArena French1386—
LMArena German1366—
LMArena Japanese1341—
LMArena Spanish1355—

Instruction Following GPT-5 Mini leads

GPT-5 Mini: 76.2 (#46), Hunyuan T1 20250711: 72.6 (#138)

Instruction Following benchmarks
BenchmarkGPT-5 MiniHunyuan T1 20250711
LMArena Instruction Following13571374
IFEval92.7%—

Long Context Too close to call

GPT-5 Mini: 41.9 (#132), Hunyuan T1 20250711: 42.2 (#128)

Long Context benchmarks
BenchmarkGPT-5 MiniHunyuan T1 20250711
LMArena Longer Query13551384
Fiction.LiveBench69.4%—

Writing & Preference Hunyuan T1 20250711 leads

GPT-5 Mini: 55.2 (#148), Hunyuan T1 20250711: 59.5 (#109)

Writing & Preference benchmarks
BenchmarkGPT-5 MiniHunyuan T1 20250711
LMArena Text13731401
LMArena Creative Writing13251392
LMArena Multi-Turn13631393
Short-Story Creative Writing83.1%—
EQ-Bench Creative Writing1313—
WildBench85.5%—

Frequently asked questions

Is GPT-5 Mini better than Hunyuan T1 20250711?

GPT-5 Mini and Hunyuan T1 20250711 score almost the same on the Noometry Index (41.8 vs 42.5), so choose on price, context window or the category you care about most.

Is GPT-5 Mini or Hunyuan T1 20250711 better for coding?

They score almost the same on coding (40.1 vs 40.9); test both on your own repository before choosing.

How many benchmarks do GPT-5 Mini and Hunyuan T1 20250711 share?

13 benchmarks have published results for both models. GPT-5 Mini has 60 scored results on Noometry and Hunyuan T1 20250711 has 13.

Related comparisons

Go deeper