Model comparison

Claude 2.1 vs Codestral

Codestral is the stronger model overall, scoring 30.6 to 25.2 on the Noometry Index.

Last verified . 0 shared benchmarks.

Claude 2.1 Anthropic

25.2

Rank #345 Reported

Codestral Mistral AI

30.6

Rank #290 Reported

Side by side

Claude 2.1 and Codestral specifications
Claude 2.1Codestral
ProviderAnthropicMistral AI
Noometry Index25.230.6
Released2023-11-212024-05-29
WeightsProprietaryProprietary
Context window—256K
Max output—8K
Input $ / M tokens—$0.30
Output $ / M tokens—$0.90
Results tracked77

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Codestral leads

Claude 2.1: 26.2 (#327), Codestral: 27.3 (#321)

Coding benchmarks
BenchmarkClaude 2.1Codestral
Aider Polyglot—11.1%
WeirdML7.1%—
BigCodeBench Instruct—41.8%
BigCodeBench Complete—52.5%
ALE-Bench—137.78
HumanEval+—73.8%
MBPP+—61.9%

Reasoning Claude 2.1 leads

Claude 2.1: 21.4 (#221), Codestral: 19.8 (#251)

Reasoning benchmarks
BenchmarkClaude 2.1Codestral
Kagi LLM Benchmark—32.5%
DTBench51%—
Epoch Capabilities Index119.27—
ForecastBench54.2—

Math Not comparable

Claude 2.1: 10.2 (#315), Codestral: —

Math benchmarks
BenchmarkClaude 2.1Codestral
OTIS Mock AIME 2024-20251.9%—

Knowledge Not comparable

Claude 2.1: 15.4 (#292), Codestral: —

Knowledge benchmarks
BenchmarkClaude 2.1Codestral
GPQA Diamond33%—
MMLU73.5%—

Frequently asked questions

Is Claude 2.1 better than Codestral?

Codestral is the stronger model overall, scoring 30.6 to 25.2 on the Noometry Index.

Is Claude 2.1 or Codestral better for coding?

Codestral scores higher on coding benchmarks: 27.3 versus 26.2 in the Noometry coding category.

How many benchmarks do Claude 2.1 and Codestral share?

0 benchmarks have published results for both models. Claude 2.1 has 7 scored results on Noometry and Codestral has 7.

Related comparisons

Go deeper