Official price table
| Tier | Input / 1M | Output / 1M | Cached input / 1M |
|---|---|---|---|
| Standard | $0.75 | $3.75 | $0.075 |
| Batch | $0.375 | $1.875 | $0.0375 |
| Flex | $0.375 | $1.875 | $0.0375 |
| Priority | $1.35 | $6.75 | $0.135 |
| Standard from Jan 1, 2027 | $1.50 | $7.50 | $0.15 |
What the promotional rate actually saves you
At launch, 3.6 Flash matched Gemini 3.5 Flash's $1.50 input and only trimmed output from $9.00 to $7.50. The promotional rate changes the shape of that: on a workload with 10M input tokens and 5M output tokens, the bill goes from $60 on 3.5 Flash to $26.25 on 3.6 Flash before caching — a 56% cut, and it lands on both sides of the ledger instead of just the output side.
Two things to hold onto. First, the discount has a date on it: budget the January 1, 2027 return to $1.50 / $7.50 rather than being surprised by it. Second, Gemini 3.7 Flash carries exactly the same price, and Google calls it the stronger model — so there is no longer a price reason to start a new project on 3.6.
Use-case fit
Use it for coding agents, grounded web workflows, document and multimodal input, and older Flash migrations where you want a stable endpoint. Use Gemini 3.5 Flash-Lite when the task is simple and price matters more than capability. Use Gemini 3.1 Pro Preview when the task is hard single-shot reasoning and you can tolerate the Pro-tier price.
Frequently asked questions
What is Gemini 3.6 Flash's API price?
Google lists Gemini 3.6 Flash at $0.75 per million input tokens, $3.75 per million output tokens, and $0.075 per million cached input tokens on the standard paid tier. Those rates hold through December 31, 2026 and double on January 1, 2027.
What are the batch and flex prices?
Batch and flex pricing are $0.375 per million input tokens and $1.875 per million output tokens. Priority pricing is $1.35 input and $6.75 output per million tokens. All of them double on January 1, 2027.
Should I use 3.6 Flash or 3.7 Flash?
They cost the same, and Google describes 3.7 Flash as the more capable model for coding, web development, and agentic work. Neither publishes a benchmark table, so the price is not the deciding factor — the provider's own positioning is.
What are the token limits?
The official model page lists a 1,048,576-token input limit and a 65,536-token output limit for gemini-3.6-flash.
Sources
- Google Gemini API pricing, verified July 22, 2026.
- Gemini 3.6 Flash model page, verified July 22, 2026.
- Gemini API release notes, verified July 22, 2026.