Model comparison

Codellama 34b Instruct vs Mistral Nemo

Codellama 34b Instruct is the stronger model overall, scoring 30.8 to 26.4 on the Noometry Index.

Last verified . 0 shared benchmarks.

Codellama 34b Instruct Meta

30.8

Rank #287 Confirmed

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Summary

  • The widest gap is in math, where Codellama 34b Instruct leads 31.0 to 25.5.

Side by side

Codellama 34b Instruct and Mistral Nemo specifications
Codellama 34b InstructMistral Nemo
ProviderMetaMistral AI
Noometry Index30.826.4
Released—2024-07-01
WeightsOpenOpen
Context window—128K
Max output—128K
Input $ / M tokens—$0.15
Output $ / M tokens—$0.15
Results tracked1410

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Codellama 34b Instruct: 28.5 (#314), Mistral Nemo: —

Coding benchmarks
BenchmarkCodellama 34b InstructMistral Nemo
BigCodeBench Instruct29%—
LMArena Coding1046—
BigCodeBench Complete37.1%—
HumanEval+43.9%—
MBPP+56.3%—

Agentic & Tool Use Not comparable

Codellama 34b Instruct: —, Mistral Nemo: 23.5 (#125)

Agentic & Tool Use benchmarks
BenchmarkCodellama 34b InstructMistral Nemo
Berkeley Function Calling Leaderboard—27.6%
BALROG—17.6%

Reasoning Mistral Nemo leads

Codellama 34b Instruct: 19.6 (#255), Mistral Nemo: 20.7 (#232)

Reasoning benchmarks
BenchmarkCodellama 34b InstructMistral Nemo
LMArena Hard Prompts1032—
DTBench—48.6%
Epoch Capabilities Index—118.68
PIQA—83.5%

Math Codellama 34b Instruct leads

Codellama 34b Instruct: 31.0 (#230), Mistral Nemo: 25.5 (#268)

Math benchmarks
BenchmarkCodellama 34b InstructMistral Nemo
LMArena Math1056—
MATH Level 5—10.8%
GSM8K—84.2%

Knowledge Not comparable

Codellama 34b Instruct: —, Mistral Nemo: 12.3 (#298)

Knowledge benchmarks
BenchmarkCodellama 34b InstructMistral Nemo
GPQA Diamond—29.9%
BoolQ—82.5%

Multilingual Not comparable

Codellama 34b Instruct: 25.8 (#284), Mistral Nemo: —

Multilingual benchmarks
BenchmarkCodellama 34b InstructMistral Nemo
LMArena Non-English1011—
LMArena Chinese976—

Instruction Following Not comparable

Codellama 34b Instruct: 52.2 (#291), Mistral Nemo: —

Instruction Following benchmarks
BenchmarkCodellama 34b InstructMistral Nemo
LMArena Instruction Following1028—

Long Context Not comparable

Codellama 34b Instruct: 30.9 (#284), Mistral Nemo: —

Long Context benchmarks
BenchmarkCodellama 34b InstructMistral Nemo
LMArena Longer Query1013—

Writing & Preference Too close to call

Codellama 34b Instruct: 28.2 (#297), Mistral Nemo: 28.5 (#296)

Writing & Preference benchmarks
BenchmarkCodellama 34b InstructMistral Nemo
LMArena Text1066—
LMArena Creative Writing1032—
EQ-Bench Creative Writing—881
LMArena Multi-Turn1015—

Frequently asked questions

Is Codellama 34b Instruct better than Mistral Nemo?

Codellama 34b Instruct is the stronger model overall, scoring 30.8 to 26.4 on the Noometry Index.

How many benchmarks do Codellama 34b Instruct and Mistral Nemo share?

0 benchmarks have published results for both models. Codellama 34b Instruct has 14 scored results on Noometry and Mistral Nemo has 10.

Related comparisons

Go deeper