model release

Arcee AI releases Trinity Large Thinking, open-source reasoning model with 262K context window

TL;DR

Arcee AI has released Trinity Large Thinking, an open-source reasoning model featuring a 262,144 token context window. The model is priced at $0.25 per million input tokens and $0.90 per million output tokens, with free access available through OpenRouter for the first five days.

1 min read
0

Arcee AI has released Trinity Large Thinking, an open-source reasoning model designed for complex problem-solving tasks. The model launched on April 1, 2026, and is available through OpenRouter with immediate access.

Specifications

Trinity Large Thinking offers a 262,144 token context window, enabling processing of substantial documents and extended conversations. The model supports OpenRouter's reasoning-enabled architecture, allowing users to access the model's step-by-step thinking process through the reasoning_details array in API responses.

Pricing

Input tokens cost $0.25 per million, while output tokens cost $0.90 per million. Arcee AI is offering free access through OpenRouter for the first five days following launch, allowing developers to evaluate the model at no cost.

Performance and Capabilities

According to Arcee AI, the model demonstrates strong performance on PinchBench, agentic workloads, and reasoning tasks. The reasoning capabilities enable transparent access to the model's internal decision-making process, allowing users to preserve and continue reasoning chains across multi-turn conversations.

Technical Details

Trinity Large Thinking is classified as a reasoning model and is available as open-source weights. OpenRouter routes requests across multiple providers with automatic fallback capability to maximize uptime. The platform normalizes requests and responses across different provider infrastructures.

Developers can integrate the model using standard OpenRouter API endpoints, with optional framework support through third-party SDKs. The reasoning parameter enables thinking mode in requests, while continuing conversations requires preservation of complete reasoning_details to maintain context continuity.

What this means

The release marks Arcee AI's entry into the reasoning model category with a competitive context window size. At $0.25/$0.90 pricing, Trinity Large Thinking positions itself in the mid-range of open-source reasoning models. The five-day free trial strategy reduces friction for adoption testing. The transparent reasoning process differentiates it from black-box alternatives, though real-world performance claims require independent verification against established benchmarks.

Related Articles

model release

Meta Releases Muse Glimmer 30B, an Open-Weight Agentic Model for Consumer Hardware

Meta Superintelligence Labs has released Muse Glimmer 30B, a dense open-weight model distilled from its larger Muse Spark system and tuned for agentic workflows on consumer hardware. The model supports 131K context, image understanding, and over 100 languages at $0.30/$1.10 per 1M input/output tokens.

model release

Fireworks Releases Ember-1, a Reasoning Model That Cuts Token Usage 40% Versus Its Kimi K3 Base

Fireworks Research has released Ember-1, a reasoning model built on Kimi K3 that produces shorter reasoning traces while claiming comparable output quality. The model offers a 1 million token context window at $3 per 1M input tokens and $15 per 1M output tokens.

model release

Z.ai Releases GLM-5.3-Prime, a High-Throughput Variant of GLM-5.3 with 1M-Token Context

Z.ai has released GLM-5.3-Prime, a high-speed variant of its GLM-5.3 model that delivers 1.5-2x the output throughput through inference acceleration while retaining the full 1M-token context window. The model is priced at $2.80 per 1M input tokens and $8.80 per 1M output tokens, targeting coding and long-horizon agentic workloads.

model release

Anonymous Stealth Model "Space Bunny Alpha" Debuts on OpenRouter With 1M-Token Context, Free During Preview

A previously unknown AI provider has released Space Bunny Alpha, a stealth model on OpenRouter offering a 1M-token context window, adjustable reasoning effort, and multimodal input support. The model is free during its preview period, though its developer remains unnamed.

Comments

Loading...