Model comparison

Codellama 34b Instruct vs Qwen3-Next 80B-A3B Instruct

Qwen3-Next 80B-A3B Instruct is the stronger model overall, scoring 43.0 to 30.8 on the Noometry Index.

Last verified . 10 shared benchmarks.

Codellama 34b Instruct Meta

30.8

Rank #287 Confirmed

Summary

  • They share 10 benchmarks with published results for both. Codellama 34b Instruct scores higher in 0 categories and Qwen3-Next 80B-A3B Instruct in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen3-Next 80B-A3B Instruct leads 58.0 to 28.2.

Side by side

Codellama 34b Instruct and Qwen3-Next 80B-A3B Instruct specifications
Codellama 34b InstructQwen3-Next 80B-A3B Instruct
ProviderMetaAlibaba (Qwen)
Noometry Index30.843.0
Released—2025-09
WeightsOpenOpen
Context window—131K
Max output—33K
Input $ / M tokens—$0.50
Output $ / M tokens—$2
Results tracked1425

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3-Next 80B-A3B Instruct leads

Codellama 34b Instruct: 28.5 (#314), Qwen3-Next 80B-A3B Instruct: 42.5 (#98)

Coding benchmarks
BenchmarkCodellama 34b InstructQwen3-Next 80B-A3B Instruct
LMArena Coding10461440
BigCodeBench Instruct29%—
BigCodeBench Complete37.1%—
HumanEval+43.9%—
MBPP+56.3%—

Reasoning Qwen3-Next 80B-A3B Instruct leads

Codellama 34b Instruct: 19.6 (#255), Qwen3-Next 80B-A3B Instruct: 31.1 (#81)

Reasoning benchmarks
BenchmarkCodellama 34b InstructQwen3-Next 80B-A3B Instruct
LMArena Hard Prompts10321428
Kagi LLM Benchmark—66.7%

Math Qwen3-Next 80B-A3B Instruct leads

Codellama 34b Instruct: 31.0 (#230), Qwen3-Next 80B-A3B Instruct: 38.8 (#126)

Math benchmarks
BenchmarkCodellama 34b InstructQwen3-Next 80B-A3B Instruct
LMArena Math10561440
Omni-MATH—46.7%

Knowledge Not comparable

Codellama 34b Instruct: —, Qwen3-Next 80B-A3B Instruct: 41.8 (#106)

Knowledge benchmarks
BenchmarkCodellama 34b InstructQwen3-Next 80B-A3B Instruct
MMLU-Pro—78.6%
Vectara Hallucination Rate—9.3%
GPQA (HELM)—63%
LMArena Expert—1417

Multilingual Qwen3-Next 80B-A3B Instruct leads

Codellama 34b Instruct: 25.8 (#284), Qwen3-Next 80B-A3B Instruct: 52.1 (#93)

Multilingual benchmarks
BenchmarkCodellama 34b InstructQwen3-Next 80B-A3B Instruct
LMArena Non-English10111407
LMArena Chinese9761460
LMArena French—1413
LMArena German—1417
LMArena Japanese—1395
LMArena Korean—1364
LMArena Russian—1404
LMArena Spanish—1435

Instruction Following Qwen3-Next 80B-A3B Instruct leads

Codellama 34b Instruct: 52.2 (#291), Qwen3-Next 80B-A3B Instruct: 70.8 (#159)

Instruction Following benchmarks
BenchmarkCodellama 34b InstructQwen3-Next 80B-A3B Instruct
LMArena Instruction Following10281389
IFEval—81%

Long Context Qwen3-Next 80B-A3B Instruct leads

Codellama 34b Instruct: 30.9 (#284), Qwen3-Next 80B-A3B Instruct: 37.0 (#223)

Long Context benchmarks
BenchmarkCodellama 34b InstructQwen3-Next 80B-A3B Instruct
LMArena Longer Query10131403
Fiction.LiveBench—55.6%

Writing & Preference Qwen3-Next 80B-A3B Instruct leads

Codellama 34b Instruct: 28.2 (#297), Qwen3-Next 80B-A3B Instruct: 58.0 (#121)

Writing & Preference benchmarks
BenchmarkCodellama 34b InstructQwen3-Next 80B-A3B Instruct
LMArena Text10661417
LMArena Creative Writing10321334
LMArena Multi-Turn10151416
WildBench—80.7%

Frequently asked questions

Is Codellama 34b Instruct better than Qwen3-Next 80B-A3B Instruct?

Qwen3-Next 80B-A3B Instruct is the stronger model overall, scoring 43.0 to 30.8 on the Noometry Index.

Is Codellama 34b Instruct or Qwen3-Next 80B-A3B Instruct better for coding?

Qwen3-Next 80B-A3B Instruct scores higher on coding benchmarks: 42.5 versus 28.5 in the Noometry coding category.

How many benchmarks do Codellama 34b Instruct and Qwen3-Next 80B-A3B Instruct share?

10 benchmarks have published results for both models. Codellama 34b Instruct has 14 scored results on Noometry and Qwen3-Next 80B-A3B Instruct has 25.

Related comparisons

Go deeper