DeepSeek-V4.1-Flash vs Qwen-Max (Qwen2.5-Max)
DeepSeek-V4.1-Flash is about 5.3× cheaper than Qwen-Max (Qwen2.5-Max) on blended token cost ($0.525 vs $2.80 per 1M). DeepSeek-V4.1-Flash (DeepSeek) is $0.525/1M blended, 1M context. Qwen-Max (Qwen2.5-Max) (Alibaba) is $2.80/1M blended, 33K context.
Last verified: 26 Sep 2026 · prices from official provider documentation
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · no spam · unsubscribe anytime.
| Spec | DeepSeek-V4.1-Flash | Qwen-Max (Qwen2.5-Max) |
|---|---|---|
| Provider | DeepSeek | Alibaba |
| Status | GA | GA |
| Input $/1M | $0.3 ◎ | $1.60 |
| Output $/1M | $1.20 ◎ | $6.40 |
| Blended $/1M | $0.525 ◎ | $2.80 |
| Context | 1M ◎ | 33K |
| Max output | 384K | 8K |
| Cutoff | — | — |
Get told when either model changes price
New models, price cuts, and deprecations — a short email when something actually changes. No spam, unsubscribe anytime.
The last few changes we caught — this is what lands in your inbox:
- DeprecatingGoogle re-pointed the Gemini 2.5 TTS migration path past the 3.1 preview and onto the new GA 3.8 TTS models
- New modelGemini 3.8 Flash TTS is GA - a flagship text-to-speech model at under half the audio rate of the preview it replaces, on a two-column price schedule
- New modelGemini 3.8 Flash-Lite TTS is GA - Google's cost-efficient TTS, which Google says was built to replace Gemini 3.1 Flash TTS
◎ You're on the watch list. We'll ping you the moment a model launches, changes price, or gets deprecated.
✕ Couldn't subscribe right now — please try again in a moment.
Free forever · powered by the same data on this page.