Model comparison

Mistral Medium 3.1 vs Phi 3 Mini 4k Instruct June 2024

Mistral Medium 3.1 and Phi 3 Mini 4k Instruct June 2024 score almost the same on the Noometry Index (31.9 vs 31.3), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Mistral Medium 3.1 Mistral AI

31.9

Rank #266 Reported

Summary

  • The widest gap is in writing & preference, where Mistral Medium 3.1 leads 55.5 to 29.6.
  • Phi 3 Mini 4k Instruct June 2024 has downloadable open weights; the other is API-only.

Side by side

Mistral Medium 3.1 and Phi 3 Mini 4k Instruct June 2024 specifications
Mistral Medium 3.1Phi 3 Mini 4k Instruct June 2024
ProviderMistral AIMicrosoft
Noometry Index31.931.3
Released——
WeightsProprietaryOpen
Context window131K—
Max output105K—
Input $ / M tokens$0.40—
Output $ / M tokens$2—
Results tracked315

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Mistral Medium 3.1: —, Phi 3 Mini 4k Instruct June 2024: 31.8 (#279)

Coding benchmarks
BenchmarkMistral Medium 3.1Phi 3 Mini 4k Instruct June 2024
LMArena Coding—1093

Reasoning Phi 3 Mini 4k Instruct June 2024 leads

Mistral Medium 3.1: 10.6 (#341), Phi 3 Mini 4k Instruct June 2024: 20.8 (#231)

Reasoning benchmarks
BenchmarkMistral Medium 3.1Phi 3 Mini 4k Instruct June 2024
NYT Connections (extended)6.5%—
Thematic Generalization20.3%—
LMArena Hard Prompts—1087

Math Not comparable

Mistral Medium 3.1: —, Phi 3 Mini 4k Instruct June 2024: 33.0 (#208)

Math benchmarks
BenchmarkMistral Medium 3.1Phi 3 Mini 4k Instruct June 2024
LMArena Math—1152

Knowledge Not comparable

Mistral Medium 3.1: —, Phi 3 Mini 4k Instruct June 2024: 28.6 (#244)

Knowledge benchmarks
BenchmarkMistral Medium 3.1Phi 3 Mini 4k Instruct June 2024
LMArena Expert—1051

Multilingual Not comparable

Mistral Medium 3.1: —, Phi 3 Mini 4k Instruct June 2024: 25.9 (#282)

Multilingual benchmarks
BenchmarkMistral Medium 3.1Phi 3 Mini 4k Instruct June 2024
LMArena Non-English—1013
LMArena Chinese—1033
LMArena German—1031
LMArena Japanese—954
LMArena Korean—880
LMArena Russian—1019

Instruction Following Not comparable

Mistral Medium 3.1: —, Phi 3 Mini 4k Instruct June 2024: 54.1 (#282)

Instruction Following benchmarks
BenchmarkMistral Medium 3.1Phi 3 Mini 4k Instruct June 2024
LMArena Instruction Following—1058

Long Context Not comparable

Mistral Medium 3.1: —, Phi 3 Mini 4k Instruct June 2024: 31.7 (#277)

Long Context benchmarks
BenchmarkMistral Medium 3.1Phi 3 Mini 4k Instruct June 2024
LMArena Longer Query—1042

Writing & Preference Mistral Medium 3.1 leads

Mistral Medium 3.1: 55.5 (#145), Phi 3 Mini 4k Instruct June 2024: 29.6 (#294)

Writing & Preference benchmarks
BenchmarkMistral Medium 3.1Phi 3 Mini 4k Instruct June 2024
LMArena Text—1080
LMArena Creative Writing—1045
EQ-Bench Creative Writing1476—
LMArena Multi-Turn—1049

Frequently asked questions

Is Mistral Medium 3.1 better than Phi 3 Mini 4k Instruct June 2024?

Mistral Medium 3.1 and Phi 3 Mini 4k Instruct June 2024 score almost the same on the Noometry Index (31.9 vs 31.3), so choose on price, context window or the category you care about most.

How many benchmarks do Mistral Medium 3.1 and Phi 3 Mini 4k Instruct June 2024 share?

0 benchmarks have published results for both models. Mistral Medium 3.1 has 3 scored results on Noometry and Phi 3 Mini 4k Instruct June 2024 has 15.

Related comparisons

Go deeper