Gemini 3.8 Flash TTS

GA

Text-to-speech model: text in, audio out. GA per Google's September 22, 2026 release note, headed 'Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS generally available (GA)', which calls it the 'flagship creative TTS model engineered for studio-grade voice fidelity, nuanced acting, regional dialects, and long-form multi-turn stability'. THE PRICE FIELDS CARRY THE COLUMN IN FORCE THROUGH 2026-12-31, not the 2027 one: Google prices this model on a two-date schedule. Paid-tier Standard per 1M tokens is $0.50 input (text) / $9.00 output (audio) / $0.125 context caching through December 31, 2026, doubling to $1.00 / $18.00 / $0.25 starting January 1, 2027. Other tiers on the same two dates: Batch $0.25 / $4.50 (caching $0.0625) then $0.50 / $9.00 (caching $0.125); Flex $0.25 / $4.50 (caching $0.025) then $0.50 / $9.00 (caching $0.05); Priority $0.90 / $16.20 (caching $0.225) then $1.80 / $32.40 (caching $0.45). Cache storage, which this schema has no field for, is $0.50 per 1M tokens per hour through 2026-12-31 and $1.00 from 2027-01-01. Free tier is 'Free of charge' on Standard and Priority and 'Not available' on Batch and Flex. Audio tokens are billed at 25 tokens per second of audio. The model page gives an 8,192-token input limit and a 16,384-token output limit, text in and audio out, 130 input languages auto-detected, caching supported, and no thinking, function calling, grounding or Live API; Batch, Flex and Priority are all supported. Default output is WAV with a RIFF header, a change from earlier Gemini TTS models that returned raw PCM. Released 2026-09-22 per both the release note and the deprecations table ('No shutdown date announced', row not gray); the model page's 'Latest update' field says July 2026, which is coarser and older, so it is not used. Google publishes no knowledge cutoff. Google's deprecations table names this model, first, as the replacement for gemini-3.1-flash-tts-preview. Added 2026-09-24, two days after launch.

Gemini 3.8 Flash TTS by Google costs $0.5 per 1M input tokens and $9 per 1M output tokens ($2.63/1M blended), with a 8K (8192-token) context window. It is generally available (GA).

Last verified: 26 Sep 2026 · sourced from official provider documentation

Provider
Google
Status
GA
Input price
$0.5 / 1M tokens
Output price
$9 / 1M tokens
Cached input
$0.125 / 1M tokens
Blended price
$2.63 / 1M tokens
Context window
8192 tokens (8K)
Max output
16.384 tokens
Modality
text, audio-out
Open weights
No — API only
Knowledge cutoff
—
Released
22 Sep 2026
API string
gemini-3.8-flash-tts

Source: Google official documentation ↗

FAQ

Gemini 3.8 Flash TTS — questions & answers

How much does Gemini 3.8 Flash TTS cost?

Gemini 3.8 Flash TTS is priced at $0.5 per 1M input tokens and $9 per 1M output tokens ($2.63/1M blended at a 3:1 input-to-output ratio), with cached input at $0.125 per 1M tokens.

What is the context window of Gemini 3.8 Flash TTS?

Gemini 3.8 Flash TTS has a 8192-token context window (8K), with up to 16.384 output tokens per request.

Is Gemini 3.8 Flash TTS deprecated?

No — Gemini 3.8 Flash TTS is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.

Does Gemini 3.8 Flash TTS support image or vision input?

Our catalog lists Gemini 3.8 Flash TTS's modalities as text, audio-out; image/vision input is not among them.

Can I self-host Gemini 3.8 Flash TTS?

No — Gemini 3.8 Flash TTS is proprietary. Its weights are not publicly released, so it is available only through Google's API (or hosted partners), not for self-hosting.

What is the API model string for Gemini 3.8 Flash TTS?

The API model identifier for Gemini 3.8 Flash TTS is "gemini-3.8-flash-tts" when calling Google's API.