Model comparison

DeepSeek-V2 (MoE-236B, May 2024) vs Mistral Nemo

Mistral Nemo has enough public results to be ranked (#337); DeepSeek-V2 (MoE-236B, May 2024) does not yet, so treat this comparison as directional.

Last verified . 2 shared benchmarks.

Mistral Nemo Mistral AI

26.4

Rank #337 Confirmed

Summary

  • They share 2 benchmarks with published results for both.

Side by side

DeepSeek-V2 (MoE-236B, May 2024) and Mistral Nemo specifications
DeepSeek-V2 (MoE-236B, May 2024)Mistral Nemo
ProviderDeepSeekMistral AI
Noometry Index40.326.4
Released2024-05-072024-07-01
WeightsOpenOpen
Context window—128K
Max output—128K
Input $ / M tokens—$0.15
Output $ / M tokens—$0.15
Results tracked1010

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

DeepSeek-V2 (MoE-236B, May 2024): 40.4 (#139), Mistral Nemo: —

Coding benchmarks
BenchmarkDeepSeek-V2 (MoE-236B, May 2024)Mistral Nemo
BigCodeBench Instruct48.9%—
BigCodeBench Complete59.4%—

Agentic & Tool Use Not comparable

DeepSeek-V2 (MoE-236B, May 2024): —, Mistral Nemo: 23.5 (#125)

Agentic & Tool Use benchmarks
BenchmarkDeepSeek-V2 (MoE-236B, May 2024)Mistral Nemo
Berkeley Function Calling Leaderboard—27.6%
BALROG—17.6%

Reasoning Not comparable

DeepSeek-V2 (MoE-236B, May 2024): —, Mistral Nemo: 20.7 (#232)

Reasoning benchmarks
BenchmarkDeepSeek-V2 (MoE-236B, May 2024)Mistral Nemo
Epoch Capabilities Index124.77118.68
PIQA83.9%83.5%
DTBench—48.6%
BIG-Bench Hard78.8%—
HellaSwag87.1%—
WinoGrande86.3%—

Math Not comparable

DeepSeek-V2 (MoE-236B, May 2024): —, Mistral Nemo: 25.5 (#268)

Math benchmarks
BenchmarkDeepSeek-V2 (MoE-236B, May 2024)Mistral Nemo
MATH Level 5—10.8%
GSM8K—84.2%

Knowledge Not comparable

DeepSeek-V2 (MoE-236B, May 2024): —, Mistral Nemo: 12.3 (#298)

Knowledge benchmarks
BenchmarkDeepSeek-V2 (MoE-236B, May 2024)Mistral Nemo
GPQA Diamond—29.9%
ARC (AI2) Challenge92.2%—
BoolQ—82.5%
MMLU78.4%—
TriviaQA80%—

Writing & Preference Not comparable

DeepSeek-V2 (MoE-236B, May 2024): —, Mistral Nemo: 28.5 (#296)

Writing & Preference benchmarks
BenchmarkDeepSeek-V2 (MoE-236B, May 2024)Mistral Nemo
EQ-Bench Creative Writing—881

Frequently asked questions

Is DeepSeek-V2 (MoE-236B, May 2024) better than Mistral Nemo?

Mistral Nemo has enough public results to be ranked (#337); DeepSeek-V2 (MoE-236B, May 2024) does not yet, so treat this comparison as directional.

How many benchmarks do DeepSeek-V2 (MoE-236B, May 2024) and Mistral Nemo share?

2 benchmarks have published results for both models. DeepSeek-V2 (MoE-236B, May 2024) has 10 scored results on Noometry and Mistral Nemo has 10.

Related comparisons

Go deeper