Model comparison

Gemini 1.5 Flash (May 2024) vs Hunyuan Large Vision

Hunyuan Large Vision is the stronger model overall, scoring 37.6 to 33.2 on the Noometry Index.

Last verified . 13 shared benchmarks.

Gemini 1.5 Flash (May 2024) Google

33.2

Rank #246 Confirmed

Hunyuan Large Vision Tencent

37.6

Rank #202 Confirmed

Summary

  • They share 13 benchmarks with published results for both. Gemini 1.5 Flash (May 2024) scores higher in 5 categories and Hunyuan Large Vision in 4 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in math, where Hunyuan Large Vision leads 35.5 to 22.1.

Side by side

Gemini 1.5 Flash (May 2024) and Hunyuan Large Vision specifications
Gemini 1.5 Flash (May 2024)Hunyuan Large Vision
ProviderGoogleTencent
Noometry Index33.237.6
Released2024-05-14—
WeightsProprietaryProprietary
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked4213

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Large Vision leads

Gemini 1.5 Flash (May 2024): 34.4 (#236), Hunyuan Large Vision: 38.3 (#180)

Coding benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Hunyuan Large Vision
LMArena Coding12611308
WeirdML24.9%—
BigCodeBench Instruct43.5%—
BigCodeBench Complete55.1%—
HumanEval+75.6%—
MBPP+67.5%—

Agentic & Tool Use Not comparable

Gemini 1.5 Flash (May 2024): 26.6 (#102), Hunyuan Large Vision: —

Agentic & Tool Use benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Hunyuan Large Vision
BALROG14.6%—

Reasoning Hunyuan Large Vision leads

Gemini 1.5 Flash (May 2024): 21.7 (#215), Hunyuan Large Vision: 24.8 (#158)

Reasoning benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Hunyuan Large Vision
LMArena Hard Prompts12571257
DTBench53.8%—
Epoch Capabilities Index129.36—
ForecastBench53.9—
PIQA87.5%—

Math Hunyuan Large Vision leads

Gemini 1.5 Flash (May 2024): 22.1 (#281), Hunyuan Large Vision: 35.5 (#183)

Math benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Hunyuan Large Vision
LMArena Math12691268
OTIS Mock AIME 2024-202516.3%—
Omni-MATH30.4%—
MATH Level 561.9%—
FrontierMath (Feb 2025 set)0%—
GSM8K82.4%—

Knowledge Hunyuan Large Vision leads

Gemini 1.5 Flash (May 2024): 26.2 (#260), Hunyuan Large Vision: 34.4 (#196)

Knowledge benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Hunyuan Large Vision
LMArena Expert12331252
GPQA Diamond47.3%—
MMLU-Pro67.8%—
GPQA (HELM)43.7%—
BoolQ85.8%—
MMLU77.9%—

Multimodal Too close to call

Gemini 1.5 Flash (May 2024): 36.0 (#81), Hunyuan Large Vision: 35.7 (#83)

Multimodal benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Hunyuan Large Vision
LMArena Vision11411180
Video-MME70.3%—
GeoBench76%—

Multilingual Gemini 1.5 Flash (May 2024) leads

Gemini 1.5 Flash (May 2024): 42.9 (#189), Hunyuan Large Vision: 39.9 (#221)

Multilingual benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Hunyuan Large Vision
LMArena Non-English12781236
LMArena Chinese12951287
LMArena Russian12881243
LMArena French1258—
LMArena German1262—
LMArena Japanese1252—
LMArena Korean1221—
LMArena Spanish1243—

Instruction Following Too close to call

Gemini 1.5 Flash (May 2024): 66.8 (#205), Hunyuan Large Vision: 65.8 (#215)

Instruction Following benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Hunyuan Large Vision
LMArena Instruction Following12581251
IFEval83.1%—

Long Context Too close to call

Gemini 1.5 Flash (May 2024): 39.0 (#187), Hunyuan Large Vision: 38.9 (#189)

Long Context benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Hunyuan Large Vision
LMArena Longer Query12841280

Writing & Preference Gemini 1.5 Flash (May 2024) leads

Gemini 1.5 Flash (May 2024): 48.7 (#196), Hunyuan Large Vision: 46.5 (#218)

Writing & Preference benchmarks
BenchmarkGemini 1.5 Flash (May 2024)Hunyuan Large Vision
LMArena Text12871263
LMArena Creative Writing12851248
LMArena Multi-Turn12531254
WildBench79.2%—

Frequently asked questions

Is Gemini 1.5 Flash (May 2024) better than Hunyuan Large Vision?

Hunyuan Large Vision is the stronger model overall, scoring 37.6 to 33.2 on the Noometry Index.

Is Gemini 1.5 Flash (May 2024) or Hunyuan Large Vision better for coding?

Hunyuan Large Vision scores higher on coding benchmarks: 38.3 versus 34.4 in the Noometry coding category.

How many benchmarks do Gemini 1.5 Flash (May 2024) and Hunyuan Large Vision share?

13 benchmarks have published results for both models. Gemini 1.5 Flash (May 2024) has 42 scored results on Noometry and Hunyuan Large Vision has 13.

Related comparisons

Go deeper