Model comparison

Hunyuan Large 2025 02 10 vs Llama 3.2 3B

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 28.9 on the Noometry Index.

Last verified . 12 shared benchmarks.

Hunyuan Large 2025 02 10 Tencent

38.6

Rank #184 Confirmed

Llama 3.2 3B Meta

28.9

Rank #321 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Hunyuan Large 2025 02 10 scores higher in 8 categories and Llama 3.2 3B in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Hunyuan Large 2025 02 10 leads 48.7 to 24.7.
  • Llama 3.2 3B has downloadable open weights; the other is API-only.

Side by side

Hunyuan Large 2025 02 10 and Llama 3.2 3B specifications
Hunyuan Large 2025 02 10Llama 3.2 3B
ProviderTencentMeta
Noometry Index38.628.9
Released—2024-09-24
WeightsProprietaryOpen
Context window—131K
Max output—118K
Input $ / M tokens—$0.05
Output $ / M tokens—$0.33
Results tracked1218

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 38.2 (#181), Llama 3.2 3B: 27.6 (#319)

Coding benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 3B
LMArena Coding13071098
BigCodeBench Instruct—23.4%
BigCodeBench Complete—28.3%

Agentic & Tool Use Not comparable

Hunyuan Large 2025 02 10: —, Llama 3.2 3B: 20.1 (#143)

Agentic & Tool Use benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 3B
Berkeley Function Calling Leaderboard—21.9%
BALROG—10.1%

Reasoning Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 25.5 (#148), Llama 3.2 3B: 21.0 (#228)

Reasoning benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 3B
LMArena Hard Prompts12861095

Math Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 35.8 (#178), Llama 3.2 3B: 32.4 (#214)

Math benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 3B
LMArena Math12811126

Knowledge Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 35.1 (#188), Llama 3.2 3B: 29.7 (#235)

Knowledge benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 3B
LMArena Expert12761090

Multilingual Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 42.0 (#200), Llama 3.2 3B: 26.2 (#281)

Multilingual benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 3B
LMArena Non-English12651019
LMArena Chinese13461017
LMArena Russian1266949
LMArena German—1056

Instruction Following Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 67.3 (#197), Llama 3.2 3B: 56.0 (#275)

Instruction Following benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 3B
LMArena Instruction Following12771089

Long Context Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 40.8 (#149), Llama 3.2 3B: 33.4 (#261)

Long Context benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 3B
LMArena Longer Query13411100

Writing & Preference Hunyuan Large 2025 02 10 leads

Hunyuan Large 2025 02 10: 48.7 (#197), Llama 3.2 3B: 24.7 (#307)

Writing & Preference benchmarks
BenchmarkHunyuan Large 2025 02 10Llama 3.2 3B
LMArena Text12881110
LMArena Creative Writing12641094
LMArena Multi-Turn12841105
EQ-Bench Creative Writing—595

Frequently asked questions

Is Hunyuan Large 2025 02 10 better than Llama 3.2 3B?

Hunyuan Large 2025 02 10 is the stronger model overall, scoring 38.6 to 28.9 on the Noometry Index.

Is Hunyuan Large 2025 02 10 or Llama 3.2 3B better for coding?

Hunyuan Large 2025 02 10 scores higher on coding benchmarks: 38.2 versus 27.6 in the Noometry coding category.

How many benchmarks do Hunyuan Large 2025 02 10 and Llama 3.2 3B share?

12 benchmarks have published results for both models. Hunyuan Large 2025 02 10 has 12 scored results on Noometry and Llama 3.2 3B has 18.

Related comparisons

Go deeper