# Gemma 3 1B vs Phi-2

> Gemma 3 1B has enough public results to be ranked (#353); Phi-2 does not yet, so treat this comparison as directional.

- Canonical page: https://noometry.com/compare/gemma-3-1b-vs-phi-2
- Last updated: 2026-10-11
- Shared benchmarks: 0

## Snapshot

| | Gemma 3 1B | Phi-2 |
|---|---|---|
| Provider | Google | Microsoft |
| Noometry Index | 21.1 | — |
| Rank | 353 | — |
| Context | — | — |
| Input $/M | — | — |
| Output $/M | — | — |
| Weights | Open | Open |

## Coding

- Gemma 3 1B: —
- Phi-2: —

| Benchmark | Gemma 3 1B | Phi-2 |
|---|---|---|
| HumanEval+ | — | 45.1% |
| MBPP+ | — | 54.2% |

## Agentic & Tool Use

- Gemma 3 1B: 13.7 (#152)
- Phi-2: —

| Benchmark | Gemma 3 1B | Phi-2 |
|---|---|---|
| Berkeley Function Calling Leaderboard | 7.2% | — |

## Reasoning

- Gemma 3 1B: 19.2 (#264)
- Phi-2: —

| Benchmark | Gemma 3 1B | Phi-2 |
|---|---|---|
| Chess Puzzles | 0% | — |
| Adversarial NLI | — | 42.5% |
| BIG-Bench Hard | — | 59.4% |
| Epoch Capabilities Index | — | 107.94 |
| HellaSwag | — | 53.6% |
| WinoGrande | — | 54.7% |

## Math

- Gemma 3 1B: 10.2 (#316)
- Phi-2: —

| Benchmark | Gemma 3 1B | Phi-2 |
|---|---|---|
| OTIS Mock AIME 2024-2025 | 1.1% | — |

## Knowledge

- Gemma 3 1B: 7.0 (#314)
- Phi-2: —

| Benchmark | Gemma 3 1B | Phi-2 |
|---|---|---|
| GPQA Diamond | 19.9% | — |
| ARC (AI2) Challenge | — | 75.9% |
| MMLU | — | 58.4% |
| OpenBookQA | — | 73.6% |
| TriviaQA | — | 45.2% |

## FAQ

### Is Gemma 3 1B better than Phi-2?

Gemma 3 1B has enough public results to be ranked (#353); Phi-2 does not yet, so treat this comparison as directional.

### How many benchmarks do Gemma 3 1B and Phi-2 share?

0 benchmarks have published results for both models. Gemma 3 1B has 4 scored results on Noometry and Phi-2 has 11.
