Google

Gemini 3.8 Flash

Google · mid-tier. Teams already on 3.7 Flash - identical price, newer model.

The published record

Every number below was read from the provider's own documentation on the date shown. benchr does not restate a figure it has not seen published.

API identifier
gemini-3.8-flash
Context window
1,048,576tokensSource
Maximum output
65,536tokensSource
Input
$0.75per 1M tokensSource
Output
$3.75per 1M tokensSource
Cached input
$0.075per 1M tokensSource
Released
September 2, 2026Source
License
Proprietary

Availability GA (stable) since September 2, 2026.

Record verified September 3, 2026

What the record says

Gemini 3.8 Flash Cyber is a specialised variant of the same foundational model, available only through Google's Fairwind Program for trusted defenders. It is not a separately priced API model and has no entry in the pricing documentation, so benchr does not carry a record for it. Google states 3.7 Flash remains fully supported.

What has moved

Entries in the change ledger that affect this model, newest first.

  1. Released

    Google released gemini-3.8-flash on September 2, 2026, its third Flash release in six weeks. Official docs list a 1,048,576-token input limit, 65,536 max output, and tunable thinking levels. Introductory pricing is $0.75/$3.75 per 1M input/output through December 31, 2026, doubling to $1.50/$7.50 on January 1, 2027; batch and flex are $0.375/$1.875, priority $1.35/$6.75, cached input $0.075. Google published HLE-Verified 54.9% and a 47.2% pass@1 patching figure in the announcement, neither of which maps to a benchmark column benchr tracks, so no benchmark value is recorded. The 3.8 Flash Cyber variant is Fairwind Program only and is not separately priced. Google states 3.7 Flash remains fully supported. Sources: blog.google/innovation-and-ai/models-and-research/gemini-models/3-8-flash-and-3-8-flash-cyber/ and ai.google.dev/gemini-api/docs/pricing (verified 2026-09-03).

    Google

The whole change ledger →

Which one to use

The rest of the family, with the two numbers that usually decide it. The current model is marked.

ModelsInput OutputContext window
Gemini 3.5 Flash-Lite$0.3$2.501,048,576
Gemini 3.6 Flash$0.75$3.751,048,576
Gemini 3.7 Flash$0.75$3.751,048,576
Gemini 3.8 Flash This page$0.75$3.751,048,576
Gemini 3.5 Flash$1.50$9.001,048,576
Gemini 3.1 Pro$2.00$12.001M

benchr's read

Worth it for

  • Teams already on 3.7 Flash - identical price, newer model
  • Agentic and multi-step coding work at Flash prices
  • 1M-token context with tunable thinking levels
  • Free-tier prototyping before a paid rollout

Look elsewhere if

  • You need the figures benchr has measured - this record is days old and unmeasured
  • You want published SWE-bench Verified or GPQA tables - Google published neither
  • You are budgeting past 2026 - the price doubles on January 1, 2027
  • You need the Cyber variant - it is Fairwind Program only, not a public API model

Prices, limits, and identifiers above are provider facts. This section is benchr's judgement about them.

What benchr has written about it

Pieces that name this model, newest first. The ones written about this model come before the ones that mention it in passing.

What is not on this page

Stated rather than filled in.

  • No first-token or tokens-per-second figure is recorded. benchr has not measured it and the provider does not publish one.