Model comparison
Inkling vs Laguna M.1
Inkling is the stronger model overall, scoring 44.1 to 32.5 on the Noometry Index.
Last verified . 2 shared benchmarks.
Summary
- They share 2 benchmarks with published results for both. Inkling scores higher in 2 categories and Laguna M.1 in 1 category; 3 gaps are clear of the uncertainty.
- The widest gap is in reasoning, where Inkling leads 40.4 to 23.1.
- Laguna M.1 accepts more context: 262K tokens versus 66K.
Side by side
| Inkling | Laguna M.1 | |
|---|---|---|
| Provider | Thinking Machines Lab | Poolside |
| Noometry Index | 44.1 | 32.5 |
| Released | 2026-07-15 | 2026-04-28 |
| Weights | Open | Open |
| Context window | 66K | 262K |
| Max output | 66K | 33K |
| Input $ / M tokens | $1.87 | — |
| Output $ / M tokens | $4.68 | — |
| Results tracked | 41 | 3 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Laguna M.1 leads
Inkling: 34.5 (#234), Laguna M.1: 36.6 (#204)
| Benchmark | Inkling | Laguna M.1 |
|---|---|---|
| LMArena WebDev | 1413 | 1349 |
| FrontierCode | 14% | — |
| FrontierSWE | 4.1% | — |
| SciCode | 47% | — |
| WeirdML | 32.3% | — |
| LMArena Coding | 1464 | — |
| ALE-Bench | 946 | — |
Agentic & Tool Use Not comparable
Inkling: 29.6 (#85), Laguna M.1: —
| Benchmark | Inkling | Laguna M.1 |
|---|---|---|
| APEX-Agents | 33.8% | — |
| τ²-bench Banking | 25% | — |
Reasoning Inkling leads
Inkling: 40.4 (#56), Laguna M.1: 23.1 (#184)
| Benchmark | Inkling | Laguna M.1 |
|---|---|---|
| ARC-AGI-2 | 36.5% | — |
| SimpleBench | 50% | — |
| ARC-AGI-1 | 79.5% | — |
| CritPt | 5.4% | — |
| Chess Puzzles | 21% | — |
| LMArena Hard Prompts | 1451 | — |
| DTBench | 87.5% | — |
| LMCA | 37.6% | — |
| Surface Evolver Bench | — | 15.6% |
| Epoch Capabilities Index | 148.54 | — |
Math Inkling leads
Inkling: 31.3 (#225), Laguna M.1: 21.1 (#283)
| Benchmark | Inkling | Laguna M.1 |
|---|---|---|
| ProofBench | 0% | 0% |
| FrontierMath (Tiers 1-3) | 33.3% | — |
| FrontierMath Tier 4 | 4.9% | — |
| OTIS Mock AIME 2024-2025 | 88.9% | — |
| LMArena Math | 1479 | — |
Knowledge Not comparable
Inkling: 55.1 (#49), Laguna M.1: —
| Benchmark | Inkling | Laguna M.1 |
|---|---|---|
| GPQA Diamond | 88.3% | — |
| SimpleQA Verified | 40.3% | — |
| LMArena Expert | 1465 | — |
Multilingual Not comparable
Inkling: 54.0 (#52), Laguna M.1: —
| Benchmark | Inkling | Laguna M.1 |
|---|---|---|
| LMArena Non-English | 1434 | — |
| LMArena Chinese | 1490 | — |
| LMArena French | 1458 | — |
| LMArena German | 1446 | — |
| LMArena Japanese | 1429 | — |
| LMArena Korean | 1404 | — |
| LMArena Russian | 1429 | — |
| LMArena Spanish | 1448 | — |
Instruction Following Not comparable
Inkling: 75.1 (#71), Laguna M.1: —
| Benchmark | Inkling | Laguna M.1 |
|---|---|---|
| LMArena Instruction Following | 1426 | — |
Long Context Not comparable
Inkling: 43.8 (#86), Laguna M.1: —
| Benchmark | Inkling | Laguna M.1 |
|---|---|---|
| LMArena Longer Query | 1434 | — |
Writing & Preference Not comparable
Inkling: 65.2 (#51), Laguna M.1: —
| Benchmark | Inkling | Laguna M.1 |
|---|---|---|
| LMArena Text | 1441 | — |
| LMArena Creative Writing | 1387 | — |
| EQ-Bench Creative Writing | 1611 | — |
| EQ-Bench 4 | 1226 | — |
| LMArena Multi-Turn | 1436 | — |
Frequently asked questions
Is Inkling better than Laguna M.1?
Inkling is the stronger model overall, scoring 44.1 to 32.5 on the Noometry Index.
Is Inkling or Laguna M.1 better for coding?
Laguna M.1 scores higher on coding benchmarks: 36.6 versus 34.5 in the Noometry coding category.
Which has the bigger context window?
Laguna M.1 does, with 262K tokens against 66K.
How many benchmarks do Inkling and Laguna M.1 share?
2 benchmarks have published results for both models. Inkling has 41 scored results on Noometry and Laguna M.1 has 3.