Model comparison

Gemma 3 27B vs Magistral Medium

Magistral Medium is the stronger model overall, scoring 35.2 to 30.8 on the Noometry Index. Gemma 3 27B costs 28× less per token, which makes it the better buy when Magistral Medium's lead doesn't matter for your workload.

Last verified . 20 shared benchmarks.

Gemma 3 27B Google

30.8

Rank #284 Confirmed

Magistral Medium Mistral AI

35.2

Rank #227 Confirmed

Summary

  • They share 20 benchmarks with published results for both. Gemma 3 27B scores higher in 4 categories and Magistral Medium in 4 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Magistral Medium leads 39.1 to 22.5.
  • The biggest single-benchmark swing is Kagi LLM Benchmark: 40.4% for Gemma 3 27B and 16.2% for Magistral Medium.
  • Gemma 3 27B is cheaper at $0.08 / $0.16 per million input/output tokens, against $2 / $5 for Magistral Medium.
  • Magistral Medium accepts more context: 262K tokens versus 131K.

Side by side

Gemma 3 27B and Magistral Medium specifications
Gemma 3 27BMagistral Medium
ProviderGoogleMistral AI
Noometry Index30.835.2
Released2025-03-112025-03-17
WeightsOpenOpen
Context window131K262K
Max output8K16K
Input $ / M tokens$0.08$2
Output $ / M tokens$0.16$5
Results tracked4322

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Medium leads

Gemma 3 27B: 22.5 (#334), Magistral Medium: 39.1 (#161)

Coding benchmarks
BenchmarkGemma 3 27BMagistral Medium
SciCode21.2%39.2%
LMArena Coding13221319
Aider Polyglot4.9%—
LiveBench Coding39.9%—

Agentic & Tool Use Not comparable

Gemma 3 27B: 25.1 (#110), Magistral Medium: —

Agentic & Tool Use benchmarks
BenchmarkGemma 3 27BMagistral Medium
Berkeley Function Calling Leaderboard29.5%—

Reasoning Gemma 3 27B leads

Gemma 3 27B: 16.7 (#301), Magistral Medium: 8.6 (#348)

Reasoning benchmarks
BenchmarkGemma 3 27BMagistral Medium
Kagi LLM Benchmark40.4%16.2%
CritPt0%0.3%
LMArena Hard Prompts13401267
ARC-AGI-2—0%
ARC-AGI-1—6.1%
Chess Puzzles0%—
LiveBench Reasoning43.8%—
DTBench52.5%—
LiveBench Data Analysis51.5%—
LMCA12.3%—
Epoch Capabilities Index130.04—
LiveBench50%—

Math Magistral Medium leads

Gemma 3 27B: 25.9 (#265), Magistral Medium: 35.1 (#189)

Math benchmarks
BenchmarkGemma 3 27BMagistral Medium
LMArena Math13121250
OTIS Mock AIME 2024-202522.5%—
LiveBench Math55.4%—
MATH Level 574%—

Knowledge Magistral Medium leads

Gemma 3 27B: 25.5 (#261), Magistral Medium: 33.5 (#202)

Knowledge benchmarks
BenchmarkGemma 3 27BMagistral Medium
LMArena Expert13041223
GPQA Diamond47.7%—
Confabulations40.3%—
Vectara Hallucination Rate7.4%—

Multimodal Not comparable

Gemma 3 27B: 32.6 (#100), Magistral Medium: —

Multimodal benchmarks
BenchmarkGemma 3 27BMagistral Medium
LMArena Vision1164—
GeoBench52%—

Multilingual Gemma 3 27B leads

Gemma 3 27B: 46.9 (#155), Magistral Medium: 39.6 (#224)

Multilingual benchmarks
BenchmarkGemma 3 27BMagistral Medium
LMArena Non-English13341232
LMArena Chinese13461227
LMArena French13681267
LMArena German13621248
LMArena Japanese12871175
LMArena Korean13081125
LMArena Russian13491224
LMArena Spanish13491271

Instruction Following Gemma 3 27B leads

Gemma 3 27B: 70.6 (#160), Magistral Medium: 66.0 (#211)

Instruction Following benchmarks
BenchmarkGemma 3 27BMagistral Medium
LMArena Instruction Following13211254
LiveBench Instruction Following74.9%—

Long Context Magistral Medium leads

Gemma 3 27B: 27.6 (#293), Magistral Medium: 39.3 (#183)

Long Context benchmarks
BenchmarkGemma 3 27BMagistral Medium
LMArena Longer Query13331295
Fiction.LiveBench33.3%—

Writing & Preference Gemma 3 27B leads

Gemma 3 27B: 52.5 (#168), Magistral Medium: 46.3 (#219)

Writing & Preference benchmarks
BenchmarkGemma 3 27BMagistral Medium
LMArena Text13581255
LMArena Creative Writing13461245
LMArena Multi-Turn13451275
Short-Story Creative Writing79.9%—
EQ-Bench Creative Writing1266—
LiveBench Language34.6%—

Frequently asked questions

Is Gemma 3 27B better than Magistral Medium?

Magistral Medium is the stronger model overall, scoring 35.2 to 30.8 on the Noometry Index. Gemma 3 27B costs 28× less per token, which makes it the better buy when Magistral Medium's lead doesn't matter for your workload.

Which is cheaper, Gemma 3 27B or Magistral Medium?

Gemma 3 27B is cheaper. It lists at $0.08 per million input tokens and $0.16 per million output tokens; Magistral Medium lists at $2 and $5.

Is Gemma 3 27B or Magistral Medium better for coding?

Magistral Medium scores higher on coding benchmarks: 39.1 versus 22.5 in the Noometry coding category.

Which has the bigger context window?

Magistral Medium does, with 262K tokens against 131K.

How many benchmarks do Gemma 3 27B and Magistral Medium share?

20 benchmarks have published results for both models. Gemma 3 27B has 43 scored results on Noometry and Magistral Medium has 22.

Related comparisons

Go deeper