deepseek
DeepSeek V4 Flash
- Input / 1M
- $0.1365
- Output / 1M
- $0.273
- Context
- 1,000,000 tokens
- Catalog
- Gateway catalog
About this model
DeepSeek V4 Flash
Use an endpoint supported by this model.
Get a key for this setup →const response = await fetch('https://api.minirouter.sh/v1/chat/completions', { method: 'POST', headers: { Authorization: `Bearer ${process.env.MINIROUTER_KEY}`, 'Content-Type': 'application/json', }, body: JSON.stringify({ model: 'deepseek/deepseek-v4-flash', messages: [{ role: 'user', content: 'Why is the sky blue?' }], max_tokens: 1024, }),}) const data = await response.json()console.log(data.choices[0].message.content)Overview
- Model type
- language
- Context window
- 1,000,000
- Maximum output
- 384,000
- Input / 1M tokens
- $0.1365
- Output / 1M tokens
- $0.273
- Released
- 2026-04-23
Prices include the MiniRouter fee. Compatibility snapshot: 2026-08-06. Input: text. Output: text.
Endpoints
/v1/chat/completions/v1/messages/v1/responsesPlatform-funded routing only. Caller BYOK credentials and OIDC-based upstream authentication are intentionally excluded; authenticate with a MiniRouter bearer key.
Captured parameters: max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning.
Structured output: Not enforced; a requested format is dropped and the answer is plain text. How it works
API
Use the setup panel above to switch between Chat Completions, Messages, and client configurations. Every snippet keeps the model ID unchanged.
https://api.minirouter.sh/v1deepseek/deepseek-v4-flashProviders
MiniRouter selects from the configured routes below. Priority is the catalog order, not a latency or availability score.
| Provider | Priority | Input / 1M | Output / 1M | Parameters | Access |
|---|---|---|---|---|---|
vercel | 1 | $0.1365 | $0.273 | max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning | MiniRouter balance |