Model comparison

Hunyuan Large 2025 02 10 vs Llama 3.2 90B

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 27.5 on the Noometry Index.

Last verified . 0 shared benchmarks.

Hunyuan Large 2025 02 10 Tencent

38.6

Rank #184 Confirmed

Llama 3.2 90B Meta

27.5

Rank #331 Confirmed

Summary

  • The widest gap is in math, where Hunyuan Large 2025 02 10 leads 35.8 to 11.1.
  • Llama 3.2 90B has downloadable open weights; the other is API-only.

Side by side

Hunyuan Large 2025 02 10 and Llama 3.2 90B specifications
Hunyuan Large 2025 02 10Llama 3.2 90B
ProviderTencentMeta
Noometry Index38.627.5
Released—2024-09-24
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked129

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Hunyuan Large 2025 02 10: 38.2 (#181), Llama 3.2 90B: —

Coding benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 90B
LMArena Coding1307—

Agentic & Tool Use Not comparable

Hunyuan Large 2025 02 10: —, Llama 3.2 90B: 30.0 (#80)

Agentic & Tool Use benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 90B
BALROG—27.3%

Reasoning Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 25.5 (#148), Llama 3.2 90B: 21.7 (#217)

Reasoning benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 90B
EnigmaEval—0.4%
LMArena Hard Prompts1286—
Epoch Capabilities Index—125.5

Math Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 35.8 (#178), Llama 3.2 90B: 11.1 (#308)

Math benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 90B
OTIS Mock AIME 2024-2025—2.6%
LMArena Math1281—
MATH Level 5—39.4%

Knowledge Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 35.1 (#188), Llama 3.2 90B: 21.7 (#274)

Knowledge benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 90B
GPQA Diamond—41%
LMArena Expert1276—
MMLU—80.3%

Multimodal Not comparable

Hunyuan Large 2025 02 10: —, Llama 3.2 90B: 25.4 (#124)

Multimodal benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 90B
LMArena Vision—1000
GeoBench—52%

Multilingual Not comparable

Hunyuan Large 2025 02 10: 42.0 (#200), Llama 3.2 90B: —

Multilingual benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 90B
LMArena Non-English1265—
LMArena Chinese1346—
LMArena Russian1266—

Instruction Following Not comparable

Hunyuan Large 2025 02 10: 67.3 (#197), Llama 3.2 90B: —

Instruction Following benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 90B
LMArena Instruction Following1277—

Long Context Not comparable

Hunyuan Large 2025 02 10: 40.8 (#149), Llama 3.2 90B: —

Long Context benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 90B
LMArena Longer Query1341—

Writing & Preference Not comparable

Hunyuan Large 2025 02 10: 48.7 (#197), Llama 3.2 90B: —

Writing & Preference benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 90B
LMArena Text1288—
LMArena Creative Writing1264—
LMArena Multi-Turn1284—

Frequently asked questions

Is Hunyuan Large 2025 02 10 better than Llama 3.2 90B?

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 27.5 on the Noometry Index.

How many benchmarks do Hunyuan Large 2025 02 10 and Llama 3.2 90B share?

0 benchmarks have published results for both models. Hunyuan Large 2025 02 10 has 12 scored results on Noometry and Llama 3.2 90B has 9.

Related comparisons

Go deeper