Model comparison

Gemma 3 27B vs Magistral Small

Gemma 3 27B and Magistral Small score almost the same on the Noometry Index (30.8 vs 30.2), so choose on price, context window or the category you care about most.

Last verified . 8 shared benchmarks.

Gemma 3 27B Google

30.8

Rank #284 Confirmed

Magistral Small Mistral AI

30.2

Rank #296 Confirmed

Summary

  • They share 8 benchmarks with published results for both. Gemma 3 27B scores higher in 1 category and Magistral Small in 3 categories; 3 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Magistral Small leads 38.4 to 22.5.
  • The biggest single-benchmark swing is Kagi LLM Benchmark: 40.4% for Gemma 3 27B and 6.3% for Magistral Small.
  • Gemma 3 27B is cheaper at $0.08 / $0.16 per million input/output tokens, against $0.50 / $1.50 for Magistral Small.
  • Gemma 3 27B accepts more context: 131K tokens versus 128K.

Side by side

Gemma 3 27B and Magistral Small specifications
Gemma 3 27BMagistral Small
ProviderGoogleMistral AI
Noometry Index30.830.2
Released2025-03-112025-06-10
WeightsOpenOpen
Context window131K128K
Max output8K40K
Input $ / M tokens$0.08$0.50
Output $ / M tokens$0.16$1.50
Results tracked4310

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Magistral Small leads

Gemma 3 27B: 22.5 (#334), Magistral Small: 38.4 (#176)

Coding benchmarks
BenchmarkGemma 3 27BMagistral Small
SciCode21.2%35.2%
Aider Polyglot4.9%—
LiveBench Coding39.9%—
LMArena Coding1322—

Agentic & Tool Use Not comparable

Gemma 3 27B: 25.1 (#110), Magistral Small: —

Agentic & Tool Use benchmarks
BenchmarkGemma 3 27BMagistral Small
Berkeley Function Calling Leaderboard29.5%—

Reasoning Gemma 3 27B leads

Gemma 3 27B: 16.7 (#301), Magistral Small: 6.8 (#350)

Reasoning benchmarks
BenchmarkGemma 3 27BMagistral Small
Kagi LLM Benchmark40.4%6.3%
CritPt0%0.3%
Chess Puzzles0%3%
DTBench52.5%61.3%
Epoch Capabilities Index130.04133.19
ARC-AGI-2—0%
ARC-AGI-1—5%
LiveBench Reasoning43.8%—
LMArena Hard Prompts1340—
LiveBench Data Analysis51.5%—
LMCA12.3%—
LiveBench50%—

Math Too close to call

Gemma 3 27B: 25.9 (#265), Magistral Small: 26.2 (#261)

Math benchmarks
BenchmarkGemma 3 27BMagistral Small
OTIS Mock AIME 2024-202522.5%30%
LiveBench Math55.4%—
LMArena Math1312—
MATH Level 574%—

Knowledge Magistral Small leads

Gemma 3 27B: 25.5 (#261), Magistral Small: 30.9 (#223)

Knowledge benchmarks
BenchmarkGemma 3 27BMagistral Small
GPQA Diamond47.7%56.1%
Confabulations40.3%—
Vectara Hallucination Rate7.4%—
LMArena Expert1304—

Multimodal Not comparable

Gemma 3 27B: 32.6 (#100), Magistral Small: —

Multimodal benchmarks
BenchmarkGemma 3 27BMagistral Small
LMArena Vision1164—
GeoBench52%—

Multilingual Not comparable

Gemma 3 27B: 46.9 (#155), Magistral Small: —

Multilingual benchmarks
BenchmarkGemma 3 27BMagistral Small
LMArena Non-English1334—
LMArena Chinese1346—
LMArena French1368—
LMArena German1362—
LMArena Japanese1287—
LMArena Korean1308—
LMArena Russian1349—
LMArena Spanish1349—

Instruction Following Not comparable

Gemma 3 27B: 70.6 (#160), Magistral Small: —

Instruction Following benchmarks
BenchmarkGemma 3 27BMagistral Small
LiveBench Instruction Following74.9%—
LMArena Instruction Following1321—

Long Context Not comparable

Gemma 3 27B: 27.6 (#293), Magistral Small: —

Long Context benchmarks
BenchmarkGemma 3 27BMagistral Small
Fiction.LiveBench33.3%—
LMArena Longer Query1333—

Writing & Preference Not comparable

Gemma 3 27B: 52.5 (#168), Magistral Small: —

Writing & Preference benchmarks
BenchmarkGemma 3 27BMagistral Small
LMArena Text1358—
LMArena Creative Writing1346—
Short-Story Creative Writing79.9%—
EQ-Bench Creative Writing1266—
LMArena Multi-Turn1345—
LiveBench Language34.6%—

Frequently asked questions

Is Gemma 3 27B better than Magistral Small?

Gemma 3 27B and Magistral Small score almost the same on the Noometry Index (30.8 vs 30.2), so choose on price, context window or the category you care about most.

Which is cheaper, Gemma 3 27B or Magistral Small?

Gemma 3 27B is cheaper. It lists at $0.08 per million input tokens and $0.16 per million output tokens; Magistral Small lists at $0.50 and $1.50.

Is Gemma 3 27B or Magistral Small better for coding?

Magistral Small scores higher on coding benchmarks: 38.4 versus 22.5 in the Noometry coding category.

Which has the bigger context window?

Gemma 3 27B does, with 131K tokens against 128K.

How many benchmarks do Gemma 3 27B and Magistral Small share?

8 benchmarks have published results for both models. Gemma 3 27B has 43 scored results on Noometry and Magistral Small has 10.

Related comparisons

Go deeper