Qwen3.6-Flash

GA

Most cost-effective current Flash model, alias = qwen3.6-flash-2026-04-16; native vision-language (multimodal text/image/video). Official International TIERED pricing: tier1 0<token<=256K: $0.25 in / $1.5 out; tier2 256K<token<=1M: $1 in / $4 out. Supports 50% batch-inference discount and context caching. Has an open-weight sibling listed: qwen3.6-35b-a3b (per official release notes, the qwen3.6-flash entry groups qwen3.6-flash / qwen3.6-flash-2026-04-16 / qwen3.6-35b-a3b). Context window 1M (tier to 1M). Max output not published on accessible official page (null). Released 2026-04-16, the date carried by the model's own dated alias qwen3.6-flash-2026-04-16 on Alibaba's pricing table. Re-read first-party on 2026-09-20: both tiers stated above are still published, the upper one as a rowspan continuation row beneath the model row rather than as a row of its own, which is why a reader that only matches rows naming the model sees the lower tier alone. Tail restored 2026-09-20: this note was cut at exactly 600 characters by an ingest truncation present since the initial commit, and the lost text is unrecoverable from git, so the ending is rewritten from a first-party re-read rather than restored (LEARNINGS #135).

Qwen3.6-Flash by Alibaba costs $0.25 per 1M input tokens and $1.50 per 1M output tokens ($0.563/1M blended), with a 1M (1.000.000-token) context window. It is generally available (GA).

Last verified: 26 Sep 2026 · sourced from official provider documentation

Provider
Alibaba
Status
GA
Input price
$0.25 / 1M tokens
Output price
$1.50 / 1M tokens
Cached input
—
Blended price
$0.563 / 1M tokens
Context window
1.000.000 tokens (1M)
Max output
—
Modality
text, image, video
Open weights
No — API only
Knowledge cutoff
—
Released
16 Apr 2026
API string
qwen3.6-flash

Source: Alibaba official documentation ↗

FAQ

Qwen3.6-Flash — questions & answers

How much does Qwen3.6-Flash cost?

Qwen3.6-Flash is priced at $0.25 per 1M input tokens and $1.50 per 1M output tokens ($0.563/1M blended at a 3:1 input-to-output ratio).

What is the context window of Qwen3.6-Flash?

Qwen3.6-Flash has a 1.000.000-token context window (1M).

Is Qwen3.6-Flash deprecated?

No — Qwen3.6-Flash is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.

Does Qwen3.6-Flash support image or vision input?

Yes — Qwen3.6-Flash accepts image input. Listed modalities: text, image, video.

Can I self-host Qwen3.6-Flash?

No — Qwen3.6-Flash is proprietary. Its weights are not publicly released, so it is available only through Alibaba's API (or hosted partners), not for self-hosting.

What is the API model string for Qwen3.6-Flash?

The API model identifier for Qwen3.6-Flash is "qwen3.6-flash" when calling Alibaba's API.