Google

Gemini 3.5 Flash

Google · mid-tier. Coding agents at speed.

The published record

Every number below was read from the provider's own documentation on the date shown. benchr does not restate a figure it has not seen published.

API identifier
gemini-3.5-flash
Context window
1,048,576tokensSource
Maximum output
65,536tokensSource
Input
$1.50per 1M tokensSource
Output
$9.00per 1M tokensSource
Cached input
$0.15per 1M tokensSource
Released
May 19, 2026Source
License
Proprietary

Availability GA (stable).

Record verified July 30, 2026

What the record says

Released at Google I/O (May 19, 2026), GA. Google's live pricing page checked 2026-07-30 lists standard $1.50/$9, batch $0.75/$4.50, cached input $0.15 plus cache-storage charges, and a free API tier. The launch blog publishes only the four agentic/coding/multimodal benchmarks above; no official SWE-bench Verified or GPQA figure is published for 3.5 Flash, so those fields are null. No universal latency figure is recorded.

What benchr has documented it doing

Capabilities in the benchr ledger that name this model. Documented means a provider says it works and benchr recorded where. Verified means benchr ran it.

1 capabilities name this model. All 1 are officially supported by the provider; benchr has run 0 of them itself.

What has moved

Entries in the change ledger that affect this model, newest first.

  1. Released

    Gemini 3.5 Flash GA since May 19, 2026 (Google I/O). $1.50/$9.00. Source: ai.google.dev/pricing

    Google

The whole change ledger →

Which one to use

The rest of the family, with the two numbers that usually decide it. The current model is marked.

ModelsInput OutputContext window
Gemini 3.5 Flash-Lite$0.3$2.501,048,576
Gemini 3.6 Flash$0.75$3.751,048,576
Gemini 3.7 Flash$0.75$3.751,048,576
Gemini 3.8 Flash$0.75$3.751,048,576
Gemini 3.5 Flash This page$1.50$9.001,048,576
Gemini 3.1 Pro$2.00$12.001M

benchr's read

Worth it for

  • Coding agents at speed
  • Parallel agent execution
  • Multimodal tasks
  • Default frontier-quality model

Look elsewhere if

  • You need the deepest single-call reasoning — use Gemini 3.1 Pro

Prices, limits, and identifiers above are provider facts. This section is benchr's judgement about them.

What benchr has written about it

Pieces that name this model, newest first. The ones written about this model come before the ones that mention it in passing.

What is not on this page

Stated rather than filled in.

  • No first-token or tokens-per-second figure is recorded. benchr has not measured it and the provider does not publish one.
  • benchr has not run any capability on this model itself. Everything in the capability section is documentation, not a test result.

Where this leads