Alibaba Cloud (Qwen)
Qwen-Flash API Pricing
International pricing, base tier (0-256K); 256K-1M tier is $0.25/$2 per 1M tokens. Alibaba recommends this as the replacement for Qwen-Turbo.
Verified against Alibaba Cloud (Qwen)’s pricing page on 2026-07-02.
View source ↗Facts
Specs
- FamilyQwen
- API IDqwen-flash
- ModalityText
- Context window—
- Max output—
- Release dateNot published
Pricing (per 1M tokens)
- Input$0.05 / 1M tokens
- Output$0.40 / 1M tokens
- Cached input—
- Cache write—
- Batch discount—
Notes
International pricing, base tier (0-256K); 256K-1M tier is $0.25/$2 per 1M tokens. Alibaba recommends this as the replacement for Qwen-Turbo.
Worked cost examples
Estimated monthly cost on three common usage shapes
| Workload | Est. monthly cost |
|---|---|
Chatbot 800 input / 200 output tokens per request, 100,000 requests/month | $12.00 |
RAG pipeline 6,000 input / 700 output tokens per request, 30,000 requests/month | $17.40 |
Batch summarization 20,000 input / 1,500 output tokens per request, 5,000 requests/month | $8.00 |
Price history
Changes we have documented since we began tracking
No price changes recorded since we began tracking.
Qwen-Flash pricing FAQ
- How much does Qwen-Flash cost per 1M tokens?
- Qwen-Flash costs $0.05 per 1M input tokens and $0.40 per 1M output tokens.
- How much does Qwen-Flash cost per request?
- A typical request of 800 input and 200 output tokens costs about $0.00012 on Qwen-Flash.
- Is Qwen-Flash cheaper than Qwen-Long?
- Qwen-Flash is cheaper than Qwen-Long on input tokens, by about 31%.
- Does Qwen-Flash have cache pricing?
- No cache discount is documented for Qwen-Flash on the provider's pricing page.