Gemini 3.6 Flash
Gemini 3.6 Flash delivers higher quality across coding, agentic workflows, and web development with reduced token consumption and fewer model calls compared to previous model iterations.
const response = await fetch('https://api.minirouter.sh/v1/chat/completions', { method: 'POST', headers: { Authorization: `Bearer ${process.env.MINIROUTER_KEY}`, 'Content-Type': 'application/json', }, body: JSON.stringify({ model: 'google/gemini-3.6-flash', messages: [{ role: 'user', content: 'Why is the sky blue?' }], }),}) const data = await response.json()console.log(data.choices[0].message.content)Overview
Live catalog values only. MiniRouter does not publish latency or throughput here until those measurements are connected to the public feed.
- Model type
- language
- Context window
- 1,000,000
- Maximum output
- 64,000
- Input / 1M tokens
- $1.575
- Output / 1M tokens
- $7.875
- Released
- 2026-07-21
Prices include the MiniRouter fee. Compatibility snapshot: 2026-08-06. Input: text, image, pdf, video. Output: text.
Endpoints
/v1/chat/completions/v1/messages/v1/responsesPlatform-funded routing only. Caller BYOK credentials and OIDC-based upstream authentication are intentionally excluded; authenticate with a MiniRouter bearer key.
Captured parameters: max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning.
API
Use the setup panel above to switch between Chat Completions, Messages, and supported client configuration drafts. Every snippet keeps the model ID unchanged.
https://api.minirouter.sh/v1google/gemini-3.6-flashProviders
MiniRouter selects from the configured routes below. Priority is the catalog order, not a latency or availability score.
| Provider | Priority | Input / 1M | Output / 1M | Parameters | Access |
|---|---|---|---|---|---|
vercel | 1 | $1.575 | $7.875 | max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning | MiniRouter balance |