Model comparison

Laguna M.1 vs Phi 3 Medium 4k Instruct

Laguna M.1 and Phi 3 Medium 4k Instruct score almost the same on the Noometry Index (32.5 vs 33.0), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Laguna M.1 Poolside

32.5

Rank #256 Reported

Phi 3 Medium 4k Instruct Microsoft

33.0

Rank #250 Confirmed

Summary

  • The widest gap is in math, where Phi 3 Medium 4k Instruct leads 33.4 to 21.1.

Side by side

Laguna M.1 and Phi 3 Medium 4k Instruct specifications
Laguna M.1Phi 3 Medium 4k Instruct
ProviderPoolsideMicrosoft
Noometry Index32.533.0
Released2026-04-28—
WeightsOpenOpen
Context window262K—
Max output33K—
Input $ / M tokens——
Output $ / M tokens——
Results tracked317

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Laguna M.1 leads

Laguna M.1: 36.6 (#204), Phi 3 Medium 4k Instruct: 32.9 (#268)

Coding benchmarks
BenchmarkLaguna M.1Phi 3 Medium 4k Instruct
LMArena WebDev1349—
LMArena Coding—1130

Reasoning Laguna M.1 leads

Laguna M.1: 23.1 (#184), Phi 3 Medium 4k Instruct: 21.7 (#214)

Reasoning benchmarks
BenchmarkLaguna M.1Phi 3 Medium 4k Instruct
LMArena Hard Prompts—1127
Surface Evolver Bench15.6%—

Math Phi 3 Medium 4k Instruct leads

Laguna M.1: 21.1 (#283), Phi 3 Medium 4k Instruct: 33.4 (#202)

Math benchmarks
BenchmarkLaguna M.1Phi 3 Medium 4k Instruct
ProofBench0%—
LMArena Math—1173

Knowledge Not comparable

Laguna M.1: —, Phi 3 Medium 4k Instruct: 30.2 (#229)

Knowledge benchmarks
BenchmarkLaguna M.1Phi 3 Medium 4k Instruct
LMArena Expert—1108

Multilingual Not comparable

Laguna M.1: —, Phi 3 Medium 4k Instruct: 30.6 (#263)

Multilingual benchmarks
BenchmarkLaguna M.1Phi 3 Medium 4k Instruct
LMArena Non-English—1094
LMArena Chinese—1108
LMArena French—1110
LMArena German—1101
LMArena Japanese—1042
LMArena Korean—954
LMArena Russian—1145
LMArena Spanish—1095

Instruction Following Not comparable

Laguna M.1: —, Phi 3 Medium 4k Instruct: 57.6 (#268)

Instruction Following benchmarks
BenchmarkLaguna M.1Phi 3 Medium 4k Instruct
LMArena Instruction Following—1113

Long Context Not comparable

Laguna M.1: —, Phi 3 Medium 4k Instruct: 34.0 (#255)

Long Context benchmarks
BenchmarkLaguna M.1Phi 3 Medium 4k Instruct
LMArena Longer Query—1120

Writing & Preference Not comparable

Laguna M.1: —, Phi 3 Medium 4k Instruct: 34.1 (#272)

Writing & Preference benchmarks
BenchmarkLaguna M.1Phi 3 Medium 4k Instruct
LMArena Text—1138
LMArena Creative Writing—1107
LMArena Multi-Turn—1088

Frequently asked questions

Is Laguna M.1 better than Phi 3 Medium 4k Instruct?

Laguna M.1 and Phi 3 Medium 4k Instruct score almost the same on the Noometry Index (32.5 vs 33.0), so choose on price, context window or the category you care about most.

Is Laguna M.1 or Phi 3 Medium 4k Instruct better for coding?

Laguna M.1 scores higher on coding benchmarks: 36.6 versus 32.9 in the Noometry coding category.

How many benchmarks do Laguna M.1 and Phi 3 Medium 4k Instruct share?

0 benchmarks have published results for both models. Laguna M.1 has 3 scored results on Noometry and Phi 3 Medium 4k Instruct has 17.

Related comparisons

Go deeper