Model comparison

Codellama 70b Instruct vs GLM-4.5V

GLM-4.5V is the stronger model overall, scoring 39.8 to 33.7 on the Noometry Index.

Last verified . 4 shared benchmarks.

Codellama 70b Instruct Meta

33.7

Rank #237 Confirmed

GLM-4.5V Z.ai (Zhipu)

39.8

Rank #158 Confirmed

Summary

  • They share 4 benchmarks with published results for both. Codellama 70b Instruct scores higher in 0 categories and GLM-4.5V in 5 categories; 5 gaps are clear of the uncertainty.
  • The widest gap is in multilingual, where GLM-4.5V leads 44.6 to 24.8.

Side by side

Codellama 70b Instruct and GLM-4.5V specifications
Codellama 70b InstructGLM-4.5V
ProviderMetaZ.ai (Zhipu)
Noometry Index33.739.8
Released—2025-08-11
WeightsOpenOpen
Context window—64K
Max output—16K
Input $ / M tokens—$0.60
Output $ / M tokens—$1.80
Results tracked715

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GLM-4.5V leads

Codellama 70b Instruct: 37.6 (#193), GLM-4.5V: 39.5 (#155)

Coding benchmarks
BenchmarkCodellama 70b InstructGLM-4.5V
BigCodeBench Instruct40.7%—
LMArena Coding—1347
BigCodeBench Complete49.6%—
HumanEval+65.9%—

Reasoning GLM-4.5V leads

Codellama 70b Instruct: 20.1 (#242), GLM-4.5V: 27.4 (#119)

Reasoning benchmarks
BenchmarkCodellama 70b InstructGLM-4.5V
LMArena Hard Prompts10521334
Kagi LLM Benchmark—59.8%

Math Not comparable

Codellama 70b Instruct: —, GLM-4.5V: 37.4 (#159)

Math benchmarks
BenchmarkCodellama 70b InstructGLM-4.5V
LMArena Math—1354

Knowledge Not comparable

Codellama 70b Instruct: —, GLM-4.5V: 37.5 (#156)

Knowledge benchmarks
BenchmarkCodellama 70b InstructGLM-4.5V
LMArena Expert—1353

Multimodal Not comparable

Codellama 70b Instruct: —, GLM-4.5V: 34.3 (#92)

Multimodal benchmarks
BenchmarkCodellama 70b InstructGLM-4.5V
LMArena Vision—1154

Multilingual GLM-4.5V leads

Codellama 70b Instruct: 24.8 (#288), GLM-4.5V: 44.6 (#177)

Multilingual benchmarks
BenchmarkCodellama 70b InstructGLM-4.5V
LMArena Non-English9921303
LMArena Chinese—1337
LMArena Russian—1298
LMArena Spanish—1336

Instruction Following GLM-4.5V leads

Codellama 70b Instruct: 51.9 (#293), GLM-4.5V: 69.2 (#175)

Instruction Following benchmarks
BenchmarkCodellama 70b InstructGLM-4.5V
LMArena Instruction Following10241311

Long Context Not comparable

Codellama 70b Instruct: —, GLM-4.5V: 39.6 (#171)

Long Context benchmarks
BenchmarkCodellama 70b InstructGLM-4.5V
LMArena Longer Query—1304

Writing & Preference GLM-4.5V leads

Codellama 70b Instruct: 33.4 (#277), GLM-4.5V: 52.5 (#170)

Writing & Preference benchmarks
BenchmarkCodellama 70b InstructGLM-4.5V
LMArena Text10571333
LMArena Creative Writing—1295
LMArena Multi-Turn—1332

Frequently asked questions

Is Codellama 70b Instruct better than GLM-4.5V?

GLM-4.5V is the stronger model overall, scoring 39.8 to 33.7 on the Noometry Index.

Is Codellama 70b Instruct or GLM-4.5V better for coding?

GLM-4.5V scores higher on coding benchmarks: 39.5 versus 37.6 in the Noometry coding category.

How many benchmarks do Codellama 70b Instruct and GLM-4.5V share?

4 benchmarks have published results for both models. Codellama 70b Instruct has 7 scored results on Noometry and GLM-4.5V has 15.

Related comparisons

Go deeper