Model comparison

Codellama 70b Instruct vs Qwen2.5-VL 72B Instruct

Codellama 70b Instruct is the stronger model overall, scoring 33.7 to 29.9 on the Noometry Index.

Last verified . 0 shared benchmarks.

Codellama 70b Instruct Meta

33.7

Rank #237 Confirmed

Side by side

Codellama 70b Instruct and Qwen2.5-VL 72B Instruct specifications
Codellama 70b InstructQwen2.5-VL 72B Instruct
ProviderMetaAlibaba (Qwen)
Noometry Index33.729.9
Released—2024-09
WeightsOpenOpen
Context window—131K
Max output—8K
Input $ / M tokens—$2.80
Output $ / M tokens—$8.40
Results tracked76

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Codellama 70b Instruct: 37.6 (#193), Qwen2.5-VL 72B Instruct: —

Coding benchmarks
BenchmarkCodellama 70b InstructQwen2.5-VL 72B Instruct
BigCodeBench Instruct40.7%—
BigCodeBench Complete49.6%—
HumanEval+65.9%—

Agentic & Tool Use Not comparable

Codellama 70b Instruct: —, Qwen2.5-VL 72B Instruct: 18.6 (#144)

Agentic & Tool Use benchmarks
BenchmarkCodellama 70b InstructQwen2.5-VL 72B Instruct
OSWorld—5%

Reasoning Too close to call

Codellama 70b Instruct: 20.1 (#242), Qwen2.5-VL 72B Instruct: 20.7 (#233)

Reasoning benchmarks
BenchmarkCodellama 70b InstructQwen2.5-VL 72B Instruct
Kagi LLM Benchmark—36%
LMArena Hard Prompts1052—

Multimodal Not comparable

Codellama 70b Instruct: —, Qwen2.5-VL 72B Instruct: 33.5 (#97)

Multimodal benchmarks
BenchmarkCodellama 70b InstructQwen2.5-VL 72B Instruct
LMArena Vision—1107
Video-MME—73.5%
GeoBench—62%
SpatialViz-Bench—33.3%

Multilingual Not comparable

Codellama 70b Instruct: 24.8 (#288), Qwen2.5-VL 72B Instruct: —

Multilingual benchmarks
BenchmarkCodellama 70b InstructQwen2.5-VL 72B Instruct
LMArena Non-English992—

Instruction Following Not comparable

Codellama 70b Instruct: 51.9 (#293), Qwen2.5-VL 72B Instruct: —

Instruction Following benchmarks
BenchmarkCodellama 70b InstructQwen2.5-VL 72B Instruct
LMArena Instruction Following1024—

Writing & Preference Not comparable

Codellama 70b Instruct: 33.4 (#277), Qwen2.5-VL 72B Instruct: —

Writing & Preference benchmarks
BenchmarkCodellama 70b InstructQwen2.5-VL 72B Instruct
LMArena Text1057—

Frequently asked questions

Is Codellama 70b Instruct better than Qwen2.5-VL 72B Instruct?

Codellama 70b Instruct is the stronger model overall, scoring 33.7 to 29.9 on the Noometry Index.

How many benchmarks do Codellama 70b Instruct and Qwen2.5-VL 72B Instruct share?

0 benchmarks have published results for both models. Codellama 70b Instruct has 7 scored results on Noometry and Qwen2.5-VL 72B Instruct has 6.

Related comparisons

Go deeper