Gemini 3.8 Flash
Google's third Flash release in six weeks, announced alongside a Cyber variant, with three thinking-effort levels, a 1M context window, and introductory pricing held to Dec 31, 2026.
Routable on MiniRouter
Use google/gemini-3.8-flash via the API →- Context
- 1M
- 64K maximum output per the Gemini API model page
- Thinking effort
- low / medium / high
- Medium is the default; Google recommends it for code and agent use
- Modalities
- Text, image, video, audio, docs in
- Text output only
- Weights
- Not released
- API access via the Gemini API and Vertex AI
Publisher pricing
Pricing below comes from the release announcement and does not represent a MiniRouter route or price.
Publisher-reported benchmarks
Results are grouped by unit. Percentage, Elo, and task-count results never share a scale.
| Benchmark | Reported result | Reported by | As of | Source |
|---|---|---|---|---|
| HLE-Verified | 54.9% | Sep 2, 2026 | Announcement ↗ |
- — Google names DeepSWE v1.1, Vals Finance Agent V2, and Harvey's legal agent benchmark in the announcement but prints no scores, so they are omitted here.
- — Google's cyber results (CyberGym, CWE-Bench) apply to the separate 3.8 Flash Cyber model, not the model MiniRouter routes.
- — Google reports the benchmark figures; MiniRouter has not reproduced them.
Independent measurements
Artificial Analysis runs its own evaluations and timing against each model's first-party API. These are the only numbers on this page not reported by the publisher.
Gemini 3.8 Flash (high) ranks #8 of 448 model families on their Intelligence Index. Their headline uses the highest-scoring effort setting; every listed setting is below.
- Intelligence Index
- 58.7
- Composite of their evaluations
- Coding Index
- 76.3
- Coding evaluations only
- Blended price
- $1.50
- USD per 1M tokens, 3:1 input to output, first-party list
- Output speed
- 298
- Median tokens per second
- First answer token
- 10s
- Median seconds, including reasoning
| Setting | Intelligence Index | Coding Index | Blended price | Output speed | First answer token |
|---|---|---|---|---|---|
| Gemini 3.8 Flash (high) ↗ | 58.7 | 76.3 | $1.50 | 298 | 10s |
| Gemini 3.8 Flash (medium) ↗ | 56.6 | 74.1 | $1.50 | 0 | 0.0s |
| Gemini 3.8 Flash (low) ↗ | 51.7 | 73.5 | $1.50 | 0 | 0.0s |
Source: Artificial Analysis ↗, read Sep 2, 2026. Measured on the model's first-party API, not a MiniRouter route.
What is not independently confirmed
- Every figure on this page is self-reported by the cited publisher. MiniRouter did not run these benchmarks.
- Benchmark harnesses are not standardised. Results sharing a name across vendors have not been confirmed to use identical runs.
- Publisher pricing can change, and introductory phases have stated end dates.
- Gemini 3.8 Flash is routable through MiniRouter today; live pricing on its model page is authoritative over the publisher figures here.
Related coverage
Other release dossiers and the current MiniRouter catalog.
Sources
- Read at Google ↗
Google announcement
Published Sep 2, 2026
- Read at Artificial Analysis ↗
Artificial Analysis model page
Published Sep 2, 2026
Editorial reference revised 2026-09-02. Catalog status comes from the generated routing snapshot.