Gemini 3.8 Flash
Google · mid-tier. Teams already on 3.7 Flash - identical price, newer model.
The published record
Every number below was read from the provider's own documentation on the date shown. benchr does not restate a figure it has not seen published.
- API identifier
gemini-3.8-flash- Context window
- 1,048,576tokensSource
- Maximum output
- 65,536tokensSource
- Input
- $0.75per 1M tokensSource
- Output
- $3.75per 1M tokensSource
- Cached input
- $0.075per 1M tokensSource
- Released
- September 2, 2026Source
- License
- Proprietary
Availability GA (stable) since September 2, 2026.
Record verified September 3, 2026
What the record says
Gemini 3.8 Flash Cyber is a specialised variant of the same foundational model, available only through Google's Fairwind Program for trusted defenders. It is not a separately priced API model and has no entry in the pricing documentation, so benchr does not carry a record for it. Google states 3.7 Flash remains fully supported.
What has moved
Entries in the change ledger that affect this model, newest first.
Released
Google released gemini-3.8-flash on September 2, 2026, its third Flash release in six weeks. Official docs list a 1,048,576-token input limit, 65,536 max output, and tunable thinking levels. Introductory pricing is $0.75/$3.75 per 1M input/output through December 31, 2026, doubling to $1.50/$7.50 on January 1, 2027; batch and flex are $0.375/$1.875, priority $1.35/$6.75, cached input $0.075. Google published HLE-Verified 54.9% and a 47.2% pass@1 patching figure in the announcement, neither of which maps to a benchmark column benchr tracks, so no benchmark value is recorded. The 3.8 Flash Cyber variant is Fairwind Program only and is not separately priced. Google states 3.7 Flash remains fully supported. Sources: blog.google/innovation-and-ai/models-and-research/gemini-models/3-8-flash-and-3-8-flash-cyber/ and ai.google.dev/gemini-api/docs/pricing (verified 2026-09-03).
Which one to use
The rest of the family, with the two numbers that usually decide it. The current model is marked.
| Models | Input | Output | Context window |
|---|---|---|---|
| Gemini 3.5 Flash-Lite | $0.3 | $2.50 | 1,048,576 |
| Gemini 3.6 Flash | $0.75 | $3.75 | 1,048,576 |
| Gemini 3.7 Flash | $0.75 | $3.75 | 1,048,576 |
| Gemini 3.8 Flash This page | $0.75 | $3.75 | 1,048,576 |
| Gemini 3.5 Flash | $1.50 | $9.00 | 1,048,576 |
| Gemini 3.1 Pro | $2.00 | $12.00 | 1M |
benchr's read
Worth it for
- Teams already on 3.7 Flash - identical price, newer model
- Agentic and multi-step coding work at Flash prices
- 1M-token context with tunable thinking levels
- Free-tier prototyping before a paid rollout
Look elsewhere if
- You need the figures benchr has measured - this record is days old and unmeasured
- You want published SWE-bench Verified or GPQA tables - Google published neither
- You are budgeting past 2026 - the price doubles on January 1, 2027
- You need the Cyber variant - it is Fairwind Program only, not a public API model
Prices, limits, and identifiers above are provider facts. This section is benchr's judgement about them.
What benchr has written about it
Pieces that name this model, newest first. The ones written about this model come before the ones that mention it in passing.
What is not on this page
Stated rather than filled in.
- No first-token or tokens-per-second figure is recorded. benchr has not measured it and the provider does not publish one.