Google Announced Sep 2, 2026

Gemini 3.8 Flash

Google's third Flash release in six weeks, announced alongside a Cyber variant, with three thinking-effort levels, a 1M context window, and introductory pricing held to Dec 31, 2026.

Context
1M
64K maximum output per the Gemini API model page
Thinking effort
low / medium / high
Medium is the default; Google recommends it for code and agent use
Modalities
Text, image, video, audio, docs in
Text output only
Weights
Not released
API access via the Gemini API and Vertex AI

Publisher pricing

Pricing below comes from the release announcement and does not represent a MiniRouter route or price.

Publisher-listed API price phases in US dollars per million tokens.
PhaseInput / 1MOutput / 1MBlended / 1MEffectiveSource
Introductory$0.75$3.75$1.50Sep 2, 2026Dec 31, 2026Google
Standard$1.50$7.50$3.00Jan 1, 2027open-endedGoogle

Publisher-reported benchmarks

Results are grouped by unit. Percentage, Elo, and task-count results never share a scale.

BenchmarkReported resultReported byAs ofSource
HLE-Verified54.9%GoogleSep 2, 2026Announcement ↗
  • Google names DeepSWE v1.1, Vals Finance Agent V2, and Harvey's legal agent benchmark in the announcement but prints no scores, so they are omitted here.
  • Google's cyber results (CyberGym, CWE-Bench) apply to the separate 3.8 Flash Cyber model, not the model MiniRouter routes.
  • Google reports the benchmark figures; MiniRouter has not reproduced them.

Independent measurements

Artificial Analysis runs its own evaluations and timing against each model's first-party API. These are the only numbers on this page not reported by the publisher.

Gemini 3.8 Flash (high) ranks #8 of 448 model families on their Intelligence Index. Their headline uses the highest-scoring effort setting; every listed setting is below.

Intelligence Index
58.7
Composite of their evaluations
Coding Index
76.3
Coding evaluations only
Blended price
$1.50
USD per 1M tokens, 3:1 input to output, first-party list
Output speed
298
Median tokens per second
First answer token
10s
Median seconds, including reasoning
SettingIntelligence IndexCoding IndexBlended priceOutput speedFirst answer token
Gemini 3.8 Flash (high)58.776.3$1.5029810s
Gemini 3.8 Flash (medium)56.674.1$1.5000.0s
Gemini 3.8 Flash (low)51.773.5$1.5000.0s

Source: Artificial Analysis, read Sep 2, 2026. Measured on the model's first-party API, not a MiniRouter route.

What is not independently confirmed

  • Every figure on this page is self-reported by the cited publisher. MiniRouter did not run these benchmarks.
  • Benchmark harnesses are not standardised. Results sharing a name across vendors have not been confirmed to use identical runs.
  • Publisher pricing can change, and introductory phases have stated end dates.
  • Gemini 3.8 Flash is routable through MiniRouter today; live pricing on its model page is authoritative over the publisher figures here.

Related coverage

Other release dossiers and the current MiniRouter catalog.

Sources

  1. Google announcement

    Published Sep 2, 2026

    Read at Google
  2. Artificial Analysis model page

    Published Sep 2, 2026

    Read at Artificial Analysis

Editorial reference revised 2026-09-02. Catalog status comes from the generated routing snapshot.