alibaba

Qwen 3.8 Flash Next

Qwen3.8-Flash-Next is Qwen’s experimental open-weight multimodal language model, pairing 125B parameters with just 6B activated for efficient reasoning and generation. With native 262K-token context extensible to 1M, vision support, and an architecture optimized for lower long-context latency, it is built for demanding agentic coding, tool use, visual reasoning, and multimodal automation.

text inputimage inputtext outputlanguage
const response = await fetch('https://api.minirouter.sh/v1/chat/completions', {  method: 'POST',  headers: {    Authorization: `Bearer ${process.env.MINIROUTER_KEY}`,    'Content-Type': 'application/json',  },  body: JSON.stringify({    model: 'alibaba/qwen3.8-flash-next',    messages: [{ role: 'user', content: 'Why is the sky blue?' }],  }),}) const data = await response.json()console.log(data.choices[0].message.content)
Read docs →

Overview

Model type
language
Context window
1,048,576
Maximum output
1,048,576
Input / 1M tokens
$0.126
Output / 1M tokens
$0.42
Released
2026-08-26

Prices include the MiniRouter fee. Compatibility snapshot: 2026-08-06. Input: text, image. Output: text.

Endpoints

/v1/chat/completions/v1/messages/v1/responses

Platform-funded routing only. Caller BYOK credentials and OIDC-based upstream authentication are intentionally excluded; authenticate with a MiniRouter bearer key.

Captured parameters: max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning.

API

Use the setup panel above to switch between Chat Completions, Messages, and client configurations. Every snippet keeps the model ID unchanged.

Base URLhttps://api.minirouter.sh/v1
Model IDalibaba/qwen3.8-flash-next

Providers

MiniRouter selects from the configured routes below. Priority is the catalog order, not a latency or availability score.

ProviderPriorityInput / 1MOutput / 1MParametersAccess
vercel
1$0.126$0.42max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoningMiniRouter balance