zai
GLM 5
GLM 5 is a frontier-class, general-purpose large language model optimized for complex systems engineering and long-horizon agentic tasks. It builds on the GLM 4.5 agent-centric lineage and is designed to support multi-step reasoning, math (including AIME-style benchmarks), advanced coding, and tool-augmented workflows, with long context support suitable for sophisticated agents and enterprise applications. Typical uses include autonomous agents for software engineering, data and systems troubleshooting, operations copilots, and high-end chat assistants that must break down complex tasks, call tools reliably, and reason over long sequences of instructions or documents.
const response = await fetch('https://api.minirouter.sh/v1/chat/completions', { method: 'POST', headers: { Authorization: `Bearer ${process.env.MINIROUTER_KEY}`, 'Content-Type': 'application/json', }, body: JSON.stringify({ model: 'zai/glm-5', messages: [{ role: 'user', content: 'Why is the sky blue?' }], }),}) const data = await response.json()console.log(data.choices[0].message.content)Overview
Live catalog values only. MiniRouter does not publish latency or throughput here until those measurements are connected to the public feed.
- Model type
- language
- Context window
- 202,800
- Maximum output
- 131,100
- Input / 1M tokens
- $1.05
- Output / 1M tokens
- $3.36
- Released
- 2026-02-12
Prices include the MiniRouter fee. Compatibility snapshot: 2026-08-06. Input: text. Output: text.
Endpoints
/v1/chat/completions/v1/messages/v1/responsesPlatform-funded routing only. Caller BYOK credentials and OIDC-based upstream authentication are intentionally excluded; authenticate with a MiniRouter bearer key.
Captured parameters: max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning.
API
Use the setup panel above to switch between Chat Completions, Messages, and supported client configuration drafts. Every snippet keeps the model ID unchanged.
https://api.minirouter.sh/v1zai/glm-5Providers
MiniRouter selects from the configured routes below. Priority is the catalog order, not a latency or availability score.
| Provider | Priority | Input / 1M | Output / 1M | Parameters | Access |
|---|---|---|---|---|---|
vercel | 1 | $1.05 | $3.36 | max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning | MiniRouter balance |