arcee-ai
Trinity Large Thinking
- Input / 1M
- $0.2625
- Output / 1M
- $0.945
- Context
- 262,100 tokens
- Catalog
- Gateway catalog
About this model
Trinity-Large-Thinking is a reasoning-optimized variant of Arcee AI's Trinity-Large family — a 398B-parameter sparse Mixture-of-Experts (MoE) model with approximately 13B active parameters per token. Built on Trinity-Large-Base and post-trained with extended chain-of-thought reasoning and agentic RL, Trinity-Large-Thinking delivers state-of-the-art performance on agentic benchmarks while maintaining strong general capabilities.
Use an endpoint supported by this model.
Get a key for this setup →const response = await fetch('https://api.minirouter.sh/v1/chat/completions', { method: 'POST', headers: { Authorization: `Bearer ${process.env.MINIROUTER_KEY}`, 'Content-Type': 'application/json', }, body: JSON.stringify({ model: 'arcee-ai/trinity-large-thinking', messages: [{ role: 'user', content: 'Why is the sky blue?' }], max_tokens: 1024, }),}) const data = await response.json()console.log(data.choices[0].message.content)Overview
- Model type
- language
- Context window
- 262,100
- Maximum output
- 80,000
- Input / 1M tokens
- $0.2625
- Output / 1M tokens
- $0.945
- Released
- 2026-04-01
Prices include the MiniRouter fee. Compatibility snapshot: 2026-08-06. Input: text. Output: text.
Endpoints
/v1/chat/completions/v1/messages/v1/responsesPlatform-funded routing only. Caller BYOK credentials and OIDC-based upstream authentication are intentionally excluded; authenticate with a MiniRouter bearer key.
Captured parameters: max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning.
Structured output: Not enforced; a requested format is dropped and the answer is plain text. How it works
API
Use the setup panel above to switch between Chat Completions, Messages, and client configurations. Every snippet keeps the model ID unchanged.
https://api.minirouter.sh/v1arcee-ai/trinity-large-thinkingProviders
MiniRouter selects from the configured routes below. Priority is the catalog order, not a latency or availability score.
| Provider | Priority | Input / 1M | Output / 1M | Parameters | Access |
|---|---|---|---|---|---|
vercel | 1 | $0.2625 | $0.945 | max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning | MiniRouter balance |