Gemini 3.8 Flash
GAGA/stable ('Stable: gemini-3.8-flash' on the model page; the September 2, 2026 release note is headed 'Gemini 3.8 Flash generally available (GA)'), described by Google as 'our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows' on both the model page and the pricing page. THE PRICE FIELDS CARRY THE COLUMN IN FORCE THROUGH 2026-12-31, not the 2027 one — Google prices this model on a two-date schedule: paid-tier Standard is $0.75 in / $3.75 out per 1M with context caching $0.075/1M through December 31, 2026, doubling to $1.50 / $7.50 / $0.15 starting January 1, 2027. Unlike gemini-3.7-flash, Google does NOT call this an introductory price anywhere on the pricing page, the model page or the release note — it publishes the schedule and no label for it, so none is asserted here. A dated ticket moves these fields on the cutover; scripts/check-google-prices.mjs resolves the schedule against the system clock and will report the row as DRIFT on 2027-01-01 without anyone remembering. Other published tiers, none of which fit the per-MTok fields, on the same two dates: Batch and Flex $0.375 in / $1.875 out (caching $0.0375) then $0.75 / $3.75 (caching $0.075); Priority $1.35 / $6.75 (caching $0.135) then $2.70 / $13.50 (caching $0.27). The per-1M-tokens-per-hour cache STORAGE fee this schema has no field for follows the same schedule: $0.50 through 2026-12-31, $1.00 from 2027-01-01. Grounding with Google Search or Maps: 5,000 requests/month free shared across all Gemini 3.x models, then $14 per 1,000. Free tier is 'Free of charge' on Standard and Priority, 'Not available' on Batch and Flex. Spec table gives 1,048,576 input / 65,536 output tokens, inputs text/image/video/audio/PDF and outputs text; thinking is supported at low, medium and high (Google notes 'minimal' returns an error); Live API, audio generation and image generation are not supported, computer use is supported in preview. Released 2026-09-02 per the API release notes; the model page's 'Latest update' field carries it more coarsely as 'September 2026'. Google publishes no knowledge cutoff on any of the three surfaces, so that stays null. Google announces no deprecation or shutdown for it — it does not appear in the deprecations table at all, so status is not inferred from any absence. Priced identically to gemini-3.7-flash and gemini-3.6-flash on all three axes during the current window.
Gemini 3.8 Flash by Google costs $0.75 per 1M input tokens and $3.75 per 1M output tokens ($1.50/1M blended), with a 1.0M (1.048.576-token) context window. It is generally available (GA).
Last verified: 26 Sep 2026 · sourced from official provider documentation
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · no spam · unsubscribe anytime.
Source: Google official documentation ↗
Compare Gemini 3.8 Flash head-to-head
$1.50 vs $10 blended /M · cross-provider
$1.50 vs $8 blended /M · cross-provider
$1.50 vs $0.525 blended /M · cross-provider
$1.50 vs $3 blended /M · cross-provider
$1.50 vs $0.75 blended /M · cross-provider
$1.50 vs $1.40 blended /M · cross-provider
Gemini 3.8 Flash — questions & answers
How much does Gemini 3.8 Flash cost?
Gemini 3.8 Flash is priced at $0.75 per 1M input tokens and $3.75 per 1M output tokens ($1.50/1M blended at a 3:1 input-to-output ratio), with cached input at $0.075 per 1M tokens.
What is the context window of Gemini 3.8 Flash?
Gemini 3.8 Flash has a 1.048.576-token context window (1.0M), with up to 65.536 output tokens per request.
Is Gemini 3.8 Flash deprecated?
No — Gemini 3.8 Flash is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.
Does Gemini 3.8 Flash support image or vision input?
Yes — Gemini 3.8 Flash accepts image input. Listed modalities: text, image, video, audio, pdf.
Can I self-host Gemini 3.8 Flash?
No — Gemini 3.8 Flash is proprietary. Its weights are not publicly released, so it is available only through Google's API (or hosted partners), not for self-hosting.
What is the API model string for Gemini 3.8 Flash?
The API model identifier for Gemini 3.8 Flash is "gemini-3.8-flash" when calling Google's API.
Track Gemini 3.8 Flash price & status changes
New models, price cuts, and deprecations — a short email when something actually changes. No spam, unsubscribe anytime.
The last few changes we caught — this is what lands in your inbox:
- DeprecatingGoogle re-pointed the Gemini 2.5 TTS migration path past the 3.1 preview and onto the new GA 3.8 TTS models
- New modelGemini 3.8 Flash TTS is GA - a flagship text-to-speech model at under half the audio rate of the preview it replaces, on a two-column price schedule
- New modelGemini 3.8 Flash-Lite TTS is GA - Google's cost-efficient TTS, which Google says was built to replace Gemini 3.1 Flash TTS
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · powered by the same data on this page.