Gemini 3.6 Flash API pricing

Google's stable July Flash model keeps Gemini 3.5 Flash's input price and cuts the output side of the bill.

By the benchr team · · Figures verified against official Google sources, July 22, 2026

Input / 1M$1.50Standard paid tier
Output / 1M$7.50Includes thinking tokens
Cached input / 1M$0.15Cache storage is extra
Context1.05M65,536 max output

Official price table

gemini-3.6-flash pricing, verified July 22, 2026
TierInput / 1MOutput / 1MCached input / 1M
Standard$1.50$7.50$0.15
Batch$0.75$3.75$0.075
Flex$0.75$3.75$0.075
Priority$2.70$13.50$0.27

When it is cheaper than 3.5 Flash

Gemini 3.6 Flash is cheaper only where output is a meaningful part of the bill. Input stays at $1.50 per million tokens, matching Gemini 3.5 Flash. Output falls from $9.00 to $7.50. On a workload with 10M input tokens and 5M output tokens, the model bill drops from $60 to $52.50 before caching. On input-heavy extraction with short outputs, the reduction is smaller.

Use-case fit

Use it for coding agents, grounded web workflows, document and multimodal input, and older Flash migrations where you want a stable endpoint. Use Gemini 3.5 Flash-Lite when the task is simple and price matters more than capability. Use Gemini 3.1 Pro Preview when the task is hard single-shot reasoning and you can tolerate the Pro-tier price.

Sources