inception

Mercury 2.5

Input / 1M
$0.042
Output / 1M
$0.1575
Context
260,000 tokens
Catalog
Gateway catalog
About this model

Mercury 2.5 is Inception’s diffusion-based reasoning model for chat, agents, and structured workflows, with tool calling, structured outputs, and a 260K context window.

text inputtext outputlanguage

Use an endpoint supported by this model.

Get a key for this setup →
const response = await fetch('https://api.minirouter.sh/v1/chat/completions', {  method: 'POST',  headers: {    Authorization: `Bearer ${process.env.MINIROUTER_KEY}`,    'Content-Type': 'application/json',  },  body: JSON.stringify({    model: 'inception/mercury-2.5',    messages: [{ role: 'user', content: 'Why is the sky blue?' }],  }),}) const data = await response.json()console.log(data.choices[0].message.content)
Read docs →

Overview

Model type
language
Context window
260,000
Maximum output
65,536
Input / 1M tokens
$0.042
Output / 1M tokens
$0.1575
Released
2026-09-08

Prices include the MiniRouter fee. Compatibility snapshot: 2026-08-06. Input: text. Output: text.

Endpoints

/v1/chat/completions/v1/messages/v1/responses

Platform-funded routing only. Caller BYOK credentials and OIDC-based upstream authentication are intentionally excluded; authenticate with a MiniRouter bearer key.

Captured parameters: max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning.

Structured output: Not enforced; a requested format is dropped and the answer is plain text. How it works

API

Use the setup panel above to switch between Chat Completions, Messages, and client configurations. Every snippet keeps the model ID unchanged.

Base URLhttps://api.minirouter.sh/v1
Model IDinception/mercury-2.5

Providers

MiniRouter selects from the configured routes below. Priority is the catalog order, not a latency or availability score.

ProviderPriorityInput / 1MOutput / 1MParametersAccess
vercel
1$0.042$0.1575max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoningMiniRouter balance