Arcee AI, open weights

# Trinity Large Thinking

> Trinity Large Thinking by Arcee AI, released April 2026. Ranked #185 of 354 with a Noometry Index of 38.6. API: $0.25 in / $0.80 out per M tokens. 262K context. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/trinity-large-thinking
- Last updated: 2026-10-10
- Title: Trinity Large Thinking Benchmarks, Price & Rank (October 2026)

Trinity Large Thinking by Arcee AI ranks 185th of 354 ranked models on the Noometry Index as of October 2026, with a score of 38.6. Its strongest category is knowledge, where it ranks 113th. API pricing starts at $0.25 per million input tokens and $0.80 per million output tokens, with a 262K-token context window.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #185 of 354
- **Index score:** 38.6
- **Evidence:** Confirmed 24 results
- **Provider:** [![](/logos/arcee.svg) Arcee AI](https://noometry.com/providers/arcee)
- **Released:** April 1, 2026
- **Weights:** Open weights
- **Reasoning:** Yes
- **Context window:** 262K
- **Max output:** 80K
- **Input price:** $0.25 / M
- **Output price:** $0.80 / M
- **Blended price:** $0.39 / M
- **Output speed:** Not measured
- **Value:** #61 of 219
- **Knowledge cutoff:** Unknown
- **Input:** text
- **Hugging Face:** [arcee-ai/Trinity-Large-Thinking](https://huggingface.co/arcee-ai/Trinity-Large-Thinking)

## Category scores

Each category score combines every public result we have in that category.

Trinity Large Thinking category scores

1.  Coding 34.1
2.  Reasoning 16.9
3.  Math 37.6
4.  Knowledge 40.9
5.  Multilingual 46.2
6.  Instruction Following 70.5
7.  Long Context 41.3
8.  Writing & Preference 53.8
9.  020406080

Trinity Large Thinking category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 34.1 | #244 | 3 |
| [Reasoning](https://noometry.com/best/reasoning) | 16.9 | #298 | 5 |
| [Math](https://noometry.com/best/math) | 37.6 | #149 | 1 |
| [Knowledge](https://noometry.com/best/knowledge) | 40.9 | #113 | 2 |
| [Multilingual](https://noometry.com/best/multilingual) | 46.2 | #160 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 70.5 | #162 | 1 |
| [Long Context](https://noometry.com/best/long-context) | 41.3 | #144 | 1 |
| [Writing & Preference](https://noometry.com/best/writing) | 53.8 | #158 | 3 |

## Strengths and weaknesses

Categories where Trinity Large Thinking places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Trinity Large Thinking: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Knowledge](https://noometry.com/best/knowledge) | 40.9 | +3.6 | #113 of 314, top 36% |
| [Math](https://noometry.com/best/math) | 37.6 | +1.1 | #149 of 327, top 46% |
| [Long Context](https://noometry.com/best/long-context) | 41.3 | +0.3 | #144 of 296, top 49% |

### Weakest categories

Trinity Large Thinking: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Reasoning](https://noometry.com/best/reasoning) | 16.9 | −6.7 | #298 of 350, top 86% |
| [Coding](https://noometry.com/best/coding) | 34.1 | −4.6 | #244 of 340, top 72% |
| [Multilingual](https://noometry.com/best/multilingual) | 46.2 | −1.2 | #160 of 297, top 54% |

## Closest competitors

The models ranked just above and below Trinity Large Thinking. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Trinity Large Thinking
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [Qwen2.5 Plus 1127](https://noometry.com/models/qwen2-5-plus) | #181 | 38.8 | — | — | [Compare](https://noometry.com/compare/qwen2-5-plus-vs-trinity-large-thinking) |
| [Qwen3.6 Flash](https://noometry.com/models/qwen3-6-flash) | #182 | 38.8 | $0.42 | — | [Compare](https://noometry.com/compare/qwen3-6-flash-vs-trinity-large-thinking) |
| [Olmo 3 32b Think](https://noometry.com/models/olmo-3-32b-think) | #183 | 38.7 | — | — | [Compare](https://noometry.com/compare/olmo-3-32b-think-vs-trinity-large-thinking) |
| [Hunyuan Large 2025 02 10](https://noometry.com/models/hunyuan-large) | #184 | 38.6 | — | — | [Compare](https://noometry.com/compare/hunyuan-large-vs-trinity-large-thinking) |
| [GPT-5.1-Codex](https://noometry.com/models/gpt-5-1-codex) | #186 | 38.6 | $3.44 | — | [Compare](https://noometry.com/compare/gpt-5-1-codex-vs-trinity-large-thinking) |
| [Sonar](https://noometry.com/models/sonar) | #187 | 38.5 | $1 | — | [Compare](https://noometry.com/compare/sonar-vs-trinity-large-thinking) |
| [MiniMax-M2.5](https://noometry.com/models/minimax-m2-5) | #188 | 38.3 | $0.52 | 49 | [Compare](https://noometry.com/compare/minimax-m2-5-vs-trinity-large-thinking) |
| [Nova Premier 1.0](https://noometry.com/models/nova-premier-1-0) | #189 | 38.3 | $5 | 9 | [Compare](https://noometry.com/compare/nova-premier-1-0-vs-trinity-large-thinking) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Trinity Large Thinking Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena WebDev](https://noometry.com/benchmarks/arena-webdev) | 1238 | #105 of 113, top 93% | thinking | [LMArena](https://lmarena.ai/leaderboard/webdev) | 2026-10-08 |
| [SciCode](https://noometry.com/benchmarks/scicode) | 36.1% | #92 of 121, top 77% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1381 | #147 of 294, top 50% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1357 |  | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Reasoning

Trinity Large Thinking Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [NYT Connections (extended)](https://noometry.com/benchmarks/nyt-connections) | 16.5% | #81 of 91, top 90% |  | [Lech Mazur benchmarks](https://github.com/lechmazur/nyt-connections) |  |
| [CritPt](https://noometry.com/benchmarks/critpt) | 0.9% | #86 of 134, top 65% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Thematic Generalization](https://noometry.com/benchmarks/thematic-generalization) | 41.6% | #20 of 23, top 87% |  | [Lech Mazur benchmarks](https://github.com/lechmazur/generalization) |  |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1350 | #160 of 297, top 54% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1342 |  | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [Surface Evolver Bench](https://noometry.com/benchmarks/surface-evolver-bench) | 15.6% | #25 of 25, top 100% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Math

Trinity Large Thinking Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1336 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1366 | #152 of 285, top 54% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Knowledge

Trinity Large Thinking Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Vectara Hallucination Rate](https://noometry.com/benchmarks/vectara-hallucination) (lower is better) | 6.9% | #27 of 96, top 29% |  | [Vectara Hallucination Leaderboard](https://github.com/vectara/hallucination-leaderboard) |  |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1352 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1360 | #148 of 273, top 55% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Multilingual

Trinity Large Thinking Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1318 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1325 | #160 of 297, top 54% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1352 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1373 | #155 of 285, top 55% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena French](https://noometry.com/benchmarks/arena-french) | 1374 | #133 of 223, top 60% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena French](https://noometry.com/benchmarks/arena-french) | 1365 |  | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1313 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1356 | #124 of 231, top 54% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Japanese](https://noometry.com/benchmarks/arena-japanese) | 1290 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Japanese](https://noometry.com/benchmarks/arena-japanese) | 1311 | #115 of 211, top 55% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Korean](https://noometry.com/benchmarks/arena-korean) | 1258 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Korean](https://noometry.com/benchmarks/arena-korean) | 1306 | #128 of 213, top 61% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1325 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1337 | #153 of 283, top 55% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Spanish](https://noometry.com/benchmarks/arena-spanish) | 1357 | #139 of 226, top 62% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Spanish](https://noometry.com/benchmarks/arena-spanish) | 1328 |  | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Trinity Large Thinking Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1334 | #155 of 298, top 53% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1325 |  | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

Trinity Large Thinking Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1355 | #150 of 291, top 52% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1318 |  | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Trinity Large Thinking Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1339 |  |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1340 | #164 of 297, top 56% | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1320 | #152 of 295, top 52% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1307 |  | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1342 | #158 of 295, top 54% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1320 |  | thinking | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

## API pricing by provider

Trinity Large Thinking API prices
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
| --- | --- | --- | --- | --- |
| [openrouter](https://openrouter.ai/arcee-ai/trinity-large-thinking) | $0.25 | $0.80 | $0.06 | 2026-10-10 |

[All Arcee AI API prices →](https://noometry.com/llm-pricing/arcee) [Estimate your cost →](https://noometry.com/tools/cost-calculator)

## Compare Trinity Large Thinking

-   [Trinity Large Thinking vs Hunyuan Large 2025 02 10](https://noometry.com/compare/hunyuan-large-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs GPT-5.1-Codex](https://noometry.com/compare/gpt-5-1-codex-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs Olmo 3 32b Think](https://noometry.com/compare/olmo-3-32b-think-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs Sonar](https://noometry.com/compare/sonar-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs Qwen3.6 Flash](https://noometry.com/compare/qwen3-6-flash-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs MiniMax-M2.5](https://noometry.com/compare/minimax-m2-5-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs Kimi K3](https://noometry.com/compare/kimi-k3-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs Grok 4.6](https://noometry.com/compare/grok-4-6-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs Qwen3.8 Max](https://noometry.com/compare/qwen3-8-max-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-trinity-large-thinking)
-   [Trinity Large Thinking vs Muse Spark 1.3](https://noometry.com/compare/muse-spark-1-3-vs-trinity-large-thinking)

## Frequently asked questions

### How good is Trinity Large Thinking?

Trinity Large Thinking by Arcee AI ranks 185th of 354 ranked models on the Noometry Index as of October 2026, with a score of 38.6. Its strongest category is knowledge, where it ranks 113th. API pricing starts at $0.25 per million input tokens and $0.80 per million output tokens, with a 262K-token context window.

### How much does Trinity Large Thinking cost?

Trinity Large Thinking costs $0.25 per million input tokens and $0.80 per million output tokens on openrouter, with cached input at $0.06.

### What is Trinity Large Thinking's context window?

Trinity Large Thinking accepts up to 262K tokens of input and can write up to 80K tokens in one response.

### Is Trinity Large Thinking open source?

Yes. Trinity Large Thinking's weights are downloadable from Hugging Face (arcee-ai/Trinity-Large-Thinking); check the license for commercial terms.

### What are Trinity Large Thinking's strengths and weaknesses?

Relative to other ranked models, Trinity Large Thinking places best in knowledge, math, long context and lowest in reasoning, coding, multilingual.

### What is Trinity Large Thinking best at?

Its best category is knowledge, where it ranks 113th on Noometry.

### Cite this page

Noometry. (2026). Trinity Large Thinking benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/trinity-large-thinking

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/trinity-large-thinking.md).
