Gemini 3.6 Flash
GAGA/stable ('Stable' on ai.google.dev/gemini-api/docs/models, described there as Google's latest Flash). Added 2026-07-31 after Google's deprecations table began naming it as the migration target for gemini-2.5-flash and the gemini-2.0-flash line. Prices are the paid-tier STANDARD rates from the official pricing page: $1.50 in / $7.50 out per 1M, context caching $0.15/1M plus a $1.00 per 1M-tokens-per-hour storage fee this schema has no field for. Other published tiers, none of which fit the per-MTok fields either: Batch and Flex $0.75/$3.75 (caching $0.075), Priority $2.70/$13.50 (caching $0.27). Grounding with Google Search or Maps: 5,000 requests/month free shared across all Gemini 3.x models, then $14 per 1,000. Spec table gives 1,048,576 input / 65,536 output tokens and inputs text, image, video, audio, PDF. CORRECTED 2026-08-02: the note here previously asserted Google publishes no release date. That was true of the two surfaces it was read from (model page, pricing page) and false about the world — Google's DEPRECATIONS table carries it as data, 'gemini-3.6-flash | July 21, 2026 | No shutdown date announced', and the API changelog entry of 2026-07-21 says the same ('Gemini 3.6 Flash and Gemini 3.5 Flash-Lite generally available'). released is now 2026-07-21. The knowledge cutoff genuinely is unpublished on all three surfaces and stays null; the model page's 'Latest update July 2026' is a docs-edit date, not a launch. Note it is CHEAPER on output than Gemini 3.5 Flash ($7.50 vs $9.00) at the same input price, so the two Flash rows are not a straight ladder.
Gemini 3.6 Flash by Google costs $1.50 per 1M input tokens and $7.50 per 1M output tokens ($3/1M blended), with a 1.0M (1.048.576-token) context window. It is generally available (GA).
Last verified: 13 Aug 2026 · sourced from official provider documentation
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · no spam · unsubscribe anytime.
Source: Google official documentation ↗
Compare Gemini 3.6 Flash head-to-head
$3 vs $10 blended /M · cross-provider
$3 vs $11.25 blended /M · cross-provider
$3 vs $3 blended /M · cross-provider
$3 vs $0.75 blended /M · cross-provider
$3 vs $1.40 blended /M · cross-provider
$3 vs $2.80 blended /M · cross-provider
Gemini 3.6 Flash — questions & answers
How much does Gemini 3.6 Flash cost?
Gemini 3.6 Flash is priced at $1.50 per 1M input tokens and $7.50 per 1M output tokens ($3/1M blended at a 3:1 input-to-output ratio), with cached input at $0.15 per 1M tokens.
What is the context window of Gemini 3.6 Flash?
Gemini 3.6 Flash has a 1.048.576-token context window (1.0M), with up to 65.536 output tokens per request.
Is Gemini 3.6 Flash deprecated?
No — Gemini 3.6 Flash is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.
Does Gemini 3.6 Flash support image or vision input?
Yes — Gemini 3.6 Flash accepts image input. Listed modalities: text, image, video, audio, pdf.
Can I self-host Gemini 3.6 Flash?
No — Gemini 3.6 Flash is proprietary. Its weights are not publicly released, so it is available only through Google's API (or hosted partners), not for self-hosting.
What is the API model string for Gemini 3.6 Flash?
The API model identifier for Gemini 3.6 Flash is "gemini-3.6-flash" when calling Google's API.
Track Gemini 3.6 Flash price & status changes
New models, price cuts, and deprecations — a short email when something actually changes. No spam, unsubscribe anytime.
The last few changes we caught — this is what lands in your inbox:
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · powered by the same data on this page.