25 signals
HOlO V1 IS LIVEone ranked AI digest a day, scored in publicREAD HOW IT WORKS →HOlO V2 STARTSyour X account, your signals, every day

Model registry / Arcee AI / Trinity / Trinity Large Thinking

Trinity Large Thinking

Reasoning-optimized 398B MoE agent model with extended thinking for long-horizon and multi-turn tool use

Line

Trinity

Weights

Open

Released

2026-04-01

Context

262K

Input

$0.25 / 1M

Output

$0.80 / 1M

Max output

262K

Coverage

not covered yet

Overview

Trinity Large Thinking is a reasoning-optimized 398B MoE agent model from Arcee AI, designed for extended thinking in long-horizon and multi-turn tool use. It is the first version of the trinity line, released on 2026-04-01. It accepts and returns text, with a context window of 262144 tokens and a maximum output of 262144 tokens. Weights are open. It supports reasoning mode and tool calling. Pricing starts at $0.25 per million input tokens and $0.80 per million output tokens, with cached input at $0.06 per million tokens. It is served by 5 hosts.

Specs

Released2026-04-01
LineArcee AI · Trinity
WeightsOpen
Context262K tokens
Max output262K tokens
Inputtext
Outputtext
ReasoningYes
Tool callingYes
Knowledge cutoffnot stated
API idtrinity-large-thinking

Pricing

Input$0.25 / 1M tokens
Cached input$0.06 / 1M tokens
Output$0.80 / 1M tokens

First version of its line in the registry.

Strengths

  • 398B MoE architecture for reasoning-optimized tasks
  • Extended thinking for long-horizon and multi-turn tool use
  • 262k token context and output windows
  • Open weights with reasoning and tool calling modes

Best for

  • Reach for it for long-horizon agentic tasks requiring extended reasoning
  • Reach for it for multi-turn tool use with 262k token context
  • Reach for it for complex text reasoning with open weights

How to access

ProviderModel id
Arceetrinity-large-thinking
Kilo Gatewaytrinity-large-thinking
NanoGPTtrinity-large-thinking
OpenRoutertrinity-large-thinking
Vercel AI Gatewaytrinity-large-thinking

5 hosts serve this model at the price above · the maker's documentation

Trinity: every version

VersionReleasedContextInput

The registry keeps the six most recent versions of a line. Older ones stay in the file and can be surfaced without a rebuild.

FAQ

What is the context window and maximum output?
Both context window and maximum output are 262144 tokens.
Does it support tool calling and reasoning?
Yes, it supports both reasoning mode and tool calling.
Are the weights open and what is the price?
Weights are open. Pricing starts at $0.25 per million input tokens and $0.80 per million output tokens, with cached input at $0.06 per million.
Open weightsReasoning262K context