google

Gemma 4 31B IT

Gemma 4 31B is engineered to tackle the most demanding enterprise workloads and complex reasoning tasks. With an expansive 256K-token context window, the 31B model can effortlessly ingest entire codebases, and massive sets of images in a single prompt.

text inputimage inputpdf inputtext outputlanguage
const response = await fetch('https://api.minirouter.sh/v1/chat/completions', {  method: 'POST',  headers: {    Authorization: `Bearer ${process.env.MINIROUTER_KEY}`,    'Content-Type': 'application/json',  },  body: JSON.stringify({    model: 'google/gemma-4-31b-it',    messages: [{ role: 'user', content: 'Why is the sky blue?' }],  }),}) const data = await response.json()console.log(data.choices[0].message.content)
Read docs →

Overview

Live catalog values only. MiniRouter does not publish latency or throughput here until those measurements are connected to the public feed.

Model type
language
Context window
262,144
Maximum output
131,072
Input / 1M tokens
$0.147
Output / 1M tokens
$0.42
Released
2026-04-02

Prices include the MiniRouter fee. Compatibility snapshot: 2026-08-06. Input: text, image, pdf. Output: text.

Endpoints

/v1/chat/completions/v1/messages/v1/responses

Platform-funded routing only. Caller BYOK credentials and OIDC-based upstream authentication are intentionally excluded; authenticate with a MiniRouter bearer key.

Captured parameters: max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoning.

API

Use the setup panel above to switch between Chat Completions, Messages, and supported client configuration drafts. Every snippet keeps the model ID unchanged.

Base URLhttps://api.minirouter.sh/v1
Model IDgoogle/gemma-4-31b-it

Providers

MiniRouter selects from the configured routes below. Priority is the catalog order, not a latency or availability score.

ProviderPriorityInput / 1MOutput / 1MParametersAccess
vercel
1$0.147$0.42max_tokens, temperature, stop, tools, tool_choice, reasoning, include_reasoningMiniRouter balance