Qwen3.5-Flash

GA

Prior-gen Flash, alias = qwen3.5-flash-2026-02-23. Official International price: 0<token<=1M $0.1 in / $0.4 out per 1M (single tier, context to 1M). Supports 50% batch-inference discount and context caching. Still listed on pricing page Jun 22, 2026. Max output not on accessible official page (null).

Qwen3.5-Flash by Alibaba costs $0.1 per 1M input tokens and $0.4 per 1M output tokens ($0.175/1M blended), with a 1M (1.000.000-token) context window. It is generally available (GA).

Last verified: 13 Aug 2026 · sourced from official provider documentation

Provider
Alibaba
Status
GA
Input price
$0.1 / 1M tokens
Output price
$0.4 / 1M tokens
Cached input
Blended price
$0.175 / 1M tokens
Context window
1.000.000 tokens (1M)
Max output
Modality
text, image
Open weights
No — API only
Knowledge cutoff
Released
23 Feb 2026
API string
qwen3.5-flash

Source: Alibaba official documentation ↗

FAQ

Qwen3.5-Flash — questions & answers

How much does Qwen3.5-Flash cost?

Qwen3.5-Flash is priced at $0.1 per 1M input tokens and $0.4 per 1M output tokens ($0.175/1M blended at a 3:1 input-to-output ratio).

What is the context window of Qwen3.5-Flash?

Qwen3.5-Flash has a 1.000.000-token context window (1M).

Is Qwen3.5-Flash deprecated?

No — Qwen3.5-Flash is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.

Does Qwen3.5-Flash support image or vision input?

Yes — Qwen3.5-Flash accepts image input. Listed modalities: text, image.

Can I self-host Qwen3.5-Flash?

No — Qwen3.5-Flash is proprietary. Its weights are not publicly released, so it is available only through Alibaba's API (or hosted partners), not for self-hosting.

What is the API model string for Qwen3.5-Flash?

The API model identifier for Qwen3.5-Flash is "qwen3.5-flash" when calling Alibaba's API.