Model comparison
Gemini 2.0 Pro vs Trinity Large Thinking
Gemini 2.0 Pro and Trinity Large Thinking score almost the same on the Noometry Index (39.1 vs 38.6), so choose on price, context window or the category you care about most.
Last verified . 0 shared benchmarks.
Summary
- The widest gap is in long context, where Trinity Large Thinking leads 41.3 to 29.2.
- Trinity Large Thinking has downloadable open weights; the other is API-only.
Side by side
| Gemini 2.0 Pro | Trinity Large Thinking | |
|---|---|---|
| Provider | Arcee AI | |
| Noometry Index | 39.1 | 38.6 |
| Released | 2025-02-05 | 2026-04-01 |
| Weights | Proprietary | Open |
| Context window | — | 262K |
| Max output | — | 80K |
| Input $ / M tokens | — | $0.25 |
| Output $ / M tokens | — | $0.80 |
| Results tracked | 14 | 24 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Gemini 2.0 Pro leads
Gemini 2.0 Pro: 37.8 (#187), Trinity Large Thinking: 34.1 (#244)
| Benchmark | Gemini 2.0 Pro | Trinity Large Thinking |
|---|---|---|
| Aider Polyglot | 35.6% | — |
| LMArena WebDev | — | 1238 |
| SciCode | — | 36.1% |
| LiveBench Coding | 63.5% | — |
| LMArena Coding | — | 1381 |
Reasoning Gemini 2.0 Pro leads
Gemini 2.0 Pro: 22.3 (#198), Trinity Large Thinking: 16.9 (#298)
| Benchmark | Gemini 2.0 Pro | Trinity Large Thinking |
|---|---|---|
| NYT Connections (extended) | — | 16.5% |
| CritPt | — | 0.9% |
| EnigmaEval | 0.7% | — |
| Thematic Generalization | — | 41.6% |
| LiveBench Reasoning | 60.1% | — |
| LMArena Hard Prompts | — | 1350 |
| LiveBench Data Analysis | 68% | — |
| Surface Evolver Bench | — | 15.6% |
| Epoch Capabilities Index | 135.06 | — |
| LiveBench | 65.1% | — |
Math Gemini 2.0 Pro leads
Gemini 2.0 Pro: 39.7 (#100), Trinity Large Thinking: 37.6 (#149)
| Benchmark | Gemini 2.0 Pro | Trinity Large Thinking |
|---|---|---|
| LiveBench Math | 71% | — |
| LMArena Math | — | 1366 |
| MATH Level 5 | 83.5% | — |
Knowledge Trinity Large Thinking leads
Gemini 2.0 Pro: 36.5 (#167), Trinity Large Thinking: 40.9 (#113)
| Benchmark | Gemini 2.0 Pro | Trinity Large Thinking |
|---|---|---|
| GPQA Diamond | 65.7% | — |
| Confabulations | 18.4% | — |
| Vectara Hallucination Rate | — | 6.9% |
| LMArena Expert | — | 1360 |
Multilingual Not comparable
Gemini 2.0 Pro: —, Trinity Large Thinking: 46.2 (#160)
| Benchmark | Gemini 2.0 Pro | Trinity Large Thinking |
|---|---|---|
| LMArena Non-English | — | 1325 |
| LMArena Chinese | — | 1373 |
| LMArena French | — | 1374 |
| LMArena German | — | 1356 |
| LMArena Japanese | — | 1311 |
| LMArena Korean | — | 1306 |
| LMArena Russian | — | 1337 |
| LMArena Spanish | — | 1357 |
Instruction Following Gemini 2.0 Pro leads
Gemini 2.0 Pro: 75.5 (#59), Trinity Large Thinking: 70.5 (#162)
| Benchmark | Gemini 2.0 Pro | Trinity Large Thinking |
|---|---|---|
| LiveBench Instruction Following | 83.4% | — |
| LMArena Instruction Following | — | 1334 |
Long Context Trinity Large Thinking leads
Gemini 2.0 Pro: 29.2 (#292), Trinity Large Thinking: 41.3 (#144)
| Benchmark | Gemini 2.0 Pro | Trinity Large Thinking |
|---|---|---|
| Fiction.LiveBench | 41.7% | — |
| LMArena Longer Query | — | 1355 |
Writing & Preference Trinity Large Thinking leads
Gemini 2.0 Pro: 52.7 (#165), Trinity Large Thinking: 53.8 (#158)
| Benchmark | Gemini 2.0 Pro | Trinity Large Thinking |
|---|---|---|
| LMArena Text | — | 1340 |
| LMArena Creative Writing | — | 1320 |
| LMArena Multi-Turn | — | 1342 |
| LiveBench Language | 44.9% | — |
Frequently asked questions
Is Gemini 2.0 Pro better than Trinity Large Thinking?
Gemini 2.0 Pro and Trinity Large Thinking score almost the same on the Noometry Index (39.1 vs 38.6), so choose on price, context window or the category you care about most.
Is Gemini 2.0 Pro or Trinity Large Thinking better for coding?
Gemini 2.0 Pro scores higher on coding benchmarks: 37.8 versus 34.1 in the Noometry coding category.
How many benchmarks do Gemini 2.0 Pro and Trinity Large Thinking share?
0 benchmarks have published results for both models. Gemini 2.0 Pro has 14 scored results on Noometry and Trinity Large Thinking has 24.