Gemini 3.6 Flash

GA

GA/stable ('Stable' on ai.google.dev/gemini-api/docs/models, described there as Google's latest Flash). Added 2026-07-31 after Google's deprecations table began naming it as the migration target for gemini-2.5-flash and the gemini-2.0-flash line. Prices are the paid-tier STANDARD rates from the official pricing page: $1.50 in / $7.50 out per 1M, context caching $0.15/1M plus a $1.00 per 1M-tokens-per-hour storage fee this schema has no field for. Other published tiers, none of which fit the per-MTok fields either: Batch and Flex $0.75/$3.75 (caching $0.075), Priority $2.70/$13.50 (caching $0.27). Grounding with Google Search or Maps: 5,000 requests/month free shared across all Gemini 3.x models, then $14 per 1,000. Spec table gives 1,048,576 input / 65,536 output tokens and inputs text, image, video, audio, PDF. CORRECTED 2026-08-02: the note here previously asserted Google publishes no release date. That was true of the two surfaces it was read from (model page, pricing page) and false about the world — Google's DEPRECATIONS table carries it as data, 'gemini-3.6-flash | July 21, 2026 | No shutdown date announced', and the API changelog entry of 2026-07-21 says the same ('Gemini 3.6 Flash and Gemini 3.5 Flash-Lite generally available'). released is now 2026-07-21. The knowledge cutoff genuinely is unpublished on all three surfaces and stays null; the model page's 'Latest update July 2026' is a docs-edit date, not a launch. Note it is CHEAPER on output than Gemini 3.5 Flash ($7.50 vs $9.00) at the same input price, so the two Flash rows are not a straight ladder.

Gemini 3.6 Flash by Google costs $1.50 per 1M input tokens and $7.50 per 1M output tokens ($3/1M blended), with a 1.0M (1.048.576-token) context window. It is generally available (GA).

Last verified: 13 Aug 2026 · sourced from official provider documentation

Provider
Google
Status
GA
Input price
$1.50 / 1M tokens
Output price
$7.50 / 1M tokens
Cached input
$0.15 / 1M tokens
Blended price
$3 / 1M tokens
Context window
1.048.576 tokens (1.0M)
Max output
65.536 tokens
Modality
text, image, video, audio, pdf
Open weights
No — API only
Knowledge cutoff
Released
21 Jul 2026
API string
gemini-3.6-flash

Source: Google official documentation ↗

FAQ

Gemini 3.6 Flash — questions & answers

How much does Gemini 3.6 Flash cost?

Gemini 3.6 Flash is priced at $1.50 per 1M input tokens and $7.50 per 1M output tokens ($3/1M blended at a 3:1 input-to-output ratio), with cached input at $0.15 per 1M tokens.

What is the context window of Gemini 3.6 Flash?

Gemini 3.6 Flash has a 1.048.576-token context window (1.0M), with up to 65.536 output tokens per request.

Is Gemini 3.6 Flash deprecated?

No — Gemini 3.6 Flash is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.

Does Gemini 3.6 Flash support image or vision input?

Yes — Gemini 3.6 Flash accepts image input. Listed modalities: text, image, video, audio, pdf.

Can I self-host Gemini 3.6 Flash?

No — Gemini 3.6 Flash is proprietary. Its weights are not publicly released, so it is available only through Google's API (or hosted partners), not for self-hosting.

What is the API model string for Gemini 3.6 Flash?

The API model identifier for Gemini 3.6 Flash is "gemini-3.6-flash" when calling Google's API.