Gemini 3.8 Live
GAGA/stable Live API model: Google lists it under "Gemini 3 / Stable" on the models index with the badge "New - Stable", and the model page's Versions row reads "Stable: gemini-3.8-live". Google describes it as the "Default Live API model for most low-latency voice agent experiences without reasoning delays". ONE SHARED PRICING TABLE COVERS THREE MODEL IDS: the pricing page heads a single Standard table with "Gemini 3.8 Live, Gemini 3.8 Live Extended Thinking, and Gemini 3.1 Flash Live Preview", so these rates are not stated for this model alone. Rates differ per input type: input $0.75 (text) / $3.00 or $0.005 per minute (audio) / $1.00 or $0.002 per minute (image/video); output $4.50 (text) / $12.00 or $0.018 per minute (audio). Values shown are the TEXT tier, matching how the catalog carries the other multi-tier Gemini rows. No context-caching rate is published and the spec table states caching is "Not supported", so the cached-input field is null as a sourced absence rather than as an unknown. Free tier is "Free of charge". Grounding with Google Search: 5,000 free requests per month shared across Gemini 3.x, then $14 per 1,000 requests. Spec table gives 131,072 input / 65,536 output tokens, inputs text/images/audio/video and outputs text and audio; audio generation, function calling, search grounding and the Live API are supported, thinking is supported as interleaved reasoning, and caching, code execution, file search, Maps grounding, image generation, structured outputs, URL context and the Batch API are not. Google names this model as the recommended replacement for four endpoints - gemini-3.1-flash-live-preview, gemini-2.5-flash-native-audio-preview-12-2025, gemini-2.0-flash-live-001 and gemini-live-2.5-flash-preview - and its model page carries a "Migrating from Gemini 3.1 Flash Live" section stating that thinking_level is not supported here, that asynchronous function calling (behavior: NON_BLOCKING) is now the default, that proactive audio is permanently enabled, and that affective dialogue has been removed from the API. Released 2026-09-15 per the Release date column of Google's deprecations table, which states "No shutdown date announced" for it; the row is NOT gray, so nothing here is inferred from an absence. Google publishes no knowledge cutoff on any surface, so that stays null. All three Google surfaces read for this row - the models index, the model page and the pricing page - are stamped "Last updated 2026-09-15 UTC". Lifecycle read first-party from Google's deprecations table on 2026-09-16.
Gemini 3.8 Live by Google costs $0.75 per 1M input tokens and $4.50 per 1M output tokens ($1.69/1M blended), with a 131K (131.072-token) context window. It is generally available (GA).
Last verified: 26 Sep 2026 · sourced from official provider documentation
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · no spam · unsubscribe anytime.
Source: Google official documentation ↗
Gemini 3.8 Live — questions & answers
How much does Gemini 3.8 Live cost?
Gemini 3.8 Live is priced at $0.75 per 1M input tokens and $4.50 per 1M output tokens ($1.69/1M blended at a 3:1 input-to-output ratio).
What is the context window of Gemini 3.8 Live?
Gemini 3.8 Live has a 131.072-token context window (131K), with up to 65.536 output tokens per request.
Is Gemini 3.8 Live deprecated?
No — Gemini 3.8 Live is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.
Does Gemini 3.8 Live support image or vision input?
Yes — Gemini 3.8 Live accepts image input. Listed modalities: text, audio, video, image, audio-out.
Can I self-host Gemini 3.8 Live?
No — Gemini 3.8 Live is proprietary. Its weights are not publicly released, so it is available only through Google's API (or hosted partners), not for self-hosting.
What is the API model string for Gemini 3.8 Live?
The API model identifier for Gemini 3.8 Live is "gemini-3.8-live" when calling Google's API.
Track Gemini 3.8 Live price & status changes
New models, price cuts, and deprecations — a short email when something actually changes. No spam, unsubscribe anytime.
The last few changes we caught — this is what lands in your inbox:
- DeprecatingGoogle re-pointed the Gemini 2.5 TTS migration path past the 3.1 preview and onto the new GA 3.8 TTS models
- New modelGemini 3.8 Flash TTS is GA - a flagship text-to-speech model at under half the audio rate of the preview it replaces, on a two-column price schedule
- New modelGemini 3.8 Flash-Lite TTS is GA - Google's cost-efficient TTS, which Google says was built to replace Gemini 3.1 Flash TTS
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · powered by the same data on this page.