Model comparison

Hunyuan Turbo 0110 vs Llama 13b

Hunyuan Turbo 0110 is the stronger model overall, scoring 39.6 to 24.4 on the Noometry Index.

Last verified . 8 shared benchmarks.

Hunyuan Turbo 0110 Tencent

39.6

Rank #163 Confirmed

Llama 13b Meta

24.4

Rank #348 Confirmed

Summary

  • They share 8 benchmarks with published results for both. Hunyuan Turbo 0110 scores higher in 6 categories and Llama 13b in 0 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Hunyuan Turbo 0110 leads 50.2 to 13.8.
  • Llama 13b has downloadable open weights; the other is API-only.

Side by side

Hunyuan Turbo 0110 and Llama 13b specifications
Hunyuan Turbo 0110Llama 13b
ProviderTencentMeta
Noometry Index39.624.4
Released—2023-02-24
WeightsProprietaryOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1121

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hunyuan Turbo 0110 leads

Hunyuan Turbo 0110: 38.6 (#171), Llama 13b: 21.4 (#337)

Coding benchmarks
BenchmarkHunyuan Turbo 0110Llama 13b
LMArena Coding1319683

Reasoning Hunyuan Turbo 0110 leads

Hunyuan Turbo 0110: 26.1 (#137), Llama 13b: 14.0 (#329)

Reasoning benchmarks
BenchmarkHunyuan Turbo 0110Llama 13b
LMArena Hard Prompts1308728
BIG-Bench Hard—37.9%
Epoch Capabilities Index—100.58
HellaSwag—79.2%
LAMBADA—75.2%
PIQA—80.1%
WinoGrande—73%

Math Hunyuan Turbo 0110 leads

Hunyuan Turbo 0110: 35.6 (#180), Llama 13b: 26.7 (#256)

Math benchmarks
BenchmarkHunyuan Turbo 0110Llama 13b
LMArena Math1273838
GSM8K—20.6%

Knowledge Not comparable

Hunyuan Turbo 0110: —, Llama 13b: —

Knowledge benchmarks
BenchmarkHunyuan Turbo 0110Llama 13b
ARC (AI2) Challenge—52.7%
BoolQ—78.7%
MMLU—47.7%
OpenBookQA—56.4%
TriviaQA—77.9%

Multimodal Not comparable

Hunyuan Turbo 0110: —, Llama 13b: —

Multimodal benchmarks
BenchmarkHunyuan Turbo 0110Llama 13b
ScienceQA—43.3%

Multilingual Hunyuan Turbo 0110 leads

Hunyuan Turbo 0110: 43.3 (#184), Llama 13b: 16.6 (#297)

Multilingual benchmarks
BenchmarkHunyuan Turbo 0110Llama 13b
LMArena Non-English1285819
LMArena Chinese1358—
LMArena Russian1305—

Instruction Following Hunyuan Turbo 0110 leads

Hunyuan Turbo 0110: 67.4 (#195), Llama 13b: 36.7 (#305)

Instruction Following benchmarks
BenchmarkHunyuan Turbo 0110Llama 13b
LMArena Instruction Following1278781

Long Context Not comparable

Hunyuan Turbo 0110: 39.8 (#168), Llama 13b: —

Long Context benchmarks
BenchmarkHunyuan Turbo 0110Llama 13b
LMArena Longer Query1309—

Writing & Preference Hunyuan Turbo 0110 leads

Hunyuan Turbo 0110: 50.2 (#183), Llama 13b: 13.8 (#312)

Writing & Preference benchmarks
BenchmarkHunyuan Turbo 0110Llama 13b
LMArena Text1311834
LMArena Creative Writing1269794
LMArena Multi-Turn1303753

Frequently asked questions

Is Hunyuan Turbo 0110 better than Llama 13b?

Hunyuan Turbo 0110 is the stronger model overall, scoring 39.6 to 24.4 on the Noometry Index.

Is Hunyuan Turbo 0110 or Llama 13b better for coding?

Hunyuan Turbo 0110 scores higher on coding benchmarks: 38.6 versus 21.4 in the Noometry coding category.

How many benchmarks do Hunyuan Turbo 0110 and Llama 13b share?

8 benchmarks have published results for both models. Hunyuan Turbo 0110 has 11 scored results on Noometry and Llama 13b has 21.

Related comparisons

Go deeper