inception
Mercury 2.5
- Input / 1M
- $0.042
- Output / 1M
- $0.1575
- Context
- 260,000 tokens
- Catalog
- Gateway catalog
About this model
Mercury 2.5 is Inception’s diffusion-based reasoning model for chat, agents, and structured workflows, with tool calling, structured outputs, and a 260K context window.
Use an endpoint supported by this model.
Get a key for this setup →const response = await fetch('https://api.minirouter.sh/v1/chat/completions', { method: 'POST', headers: { Authorization: `Bearer ${process.env.MINIROUTER_KEY}`, 'Content-Type': 'application/json', }, body: JSON.stringify({ model: 'inception/mercury-2.5', messages: [{ role: 'user', content: 'Why is the sky blue?' }], }),}) const data = await response.json()console.log(data.choices[0].message.content)Overview
- Model type
- language
- Context window
- 260,000
- Maximum output
- 65,536
- Input / 1M tokens
- $0.042
- Output / 1M tokens
- $0.1575
- Released
- 2026-09-08
Prices include the MiniRouter fee. Compatibility snapshot: 2026-08-06. Input: text. Output: text.
Endpoints
/v1/chat/completions/v1/messages/v1/responsesPlatform-funded routing only. Caller BYOK credentials and OIDC-based upstream authentication are intentionally excluded; authenticate with a MiniRouter bearer key.
Captured parameters: max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning.
Structured output: Not enforced; a requested format is dropped and the answer is plain text. How it works
API
Use the setup panel above to switch between Chat Completions, Messages, and client configurations. Every snippet keeps the model ID unchanged.
https://api.minirouter.sh/v1inception/mercury-2.5Providers
MiniRouter selects from the configured routes below. Priority is the catalog order, not a latency or availability score.
| Provider | Priority | Input / 1M | Output / 1M | Parameters | Access |
|---|---|---|---|---|---|
vercel | 1 | $0.042 | $0.1575 | max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning | MiniRouter balance |