Qwen3.6-Flash
GAMost cost-effective current Flash model, alias = qwen3.6-flash-2026-04-16; native vision-language (multimodal text/image/video). Official International TIERED pricing: tier1 0<token<=256K: $0.25 in / $1.5 out; tier2 256K<token<=1M: $1 in / $4 out. Supports 50% batch-inference discount and context caching. Has an open-weight sibling listed: qwen3.6-35b-a3b (per official release notes, the qwen3.6-flash entry groups qwen3.6-flash / qwen3.6-flash-2026-04-16 / qwen3.6-35b-a3b). Context window 1M (tier to 1M). Max output not published on accessible official page (null). Released 2026-04-16, the date carried by the model's own dated alias qwen3.6-flash-2026-04-16 on Alibaba's pricing table. Re-read first-party on 2026-09-20: both tiers stated above are still published, the upper one as a rowspan continuation row beneath the model row rather than as a row of its own, which is why a reader that only matches rows naming the model sees the lower tier alone. Tail restored 2026-09-20: this note was cut at exactly 600 characters by an ingest truncation present since the initial commit, and the lost text is unrecoverable from git, so the ending is rewritten from a first-party re-read rather than restored (LEARNINGS #135).
Qwen3.6-Flash by Alibaba costs $0.25 per 1M input tokens and $1.50 per 1M output tokens ($0.563/1M blended), with a 1M (1.000.000-token) context window. It is generally available (GA).
Last verified: 26 Sep 2026 · sourced from official provider documentation
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · no spam · unsubscribe anytime.
Source: Alibaba official documentation ↗
Qwen3.6-Flash — questions & answers
How much does Qwen3.6-Flash cost?
Qwen3.6-Flash is priced at $0.25 per 1M input tokens and $1.50 per 1M output tokens ($0.563/1M blended at a 3:1 input-to-output ratio).
What is the context window of Qwen3.6-Flash?
Qwen3.6-Flash has a 1.000.000-token context window (1M).
Is Qwen3.6-Flash deprecated?
No — Qwen3.6-Flash is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.
Does Qwen3.6-Flash support image or vision input?
Yes — Qwen3.6-Flash accepts image input. Listed modalities: text, image, video.
Can I self-host Qwen3.6-Flash?
No — Qwen3.6-Flash is proprietary. Its weights are not publicly released, so it is available only through Alibaba's API (or hosted partners), not for self-hosting.
What is the API model string for Qwen3.6-Flash?
The API model identifier for Qwen3.6-Flash is "qwen3.6-flash" when calling Alibaba's API.
Track Qwen3.6-Flash price & status changes
New models, price cuts, and deprecations — a short email when something actually changes. No spam, unsubscribe anytime.
The last few changes we caught — this is what lands in your inbox:
- DeprecatingGoogle re-pointed the Gemini 2.5 TTS migration path past the 3.1 preview and onto the new GA 3.8 TTS models
- New modelGemini 3.8 Flash TTS is GA - a flagship text-to-speech model at under half the audio rate of the preview it replaces, on a two-column price schedule
- New modelGemini 3.8 Flash-Lite TTS is GA - Google's cost-efficient TTS, which Google says was built to replace Gemini 3.1 Flash TTS
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · powered by the same data on this page.