Alibaba Cloud (Qwen)

Qwen-Flash API Pricing

International pricing, base tier (0-256K); 256K-1M tier is $0.25/$2 per 1M tokens. Alibaba recommends this as the replacement for Qwen-Turbo.

Verified against Alibaba Cloud (Qwen)’s pricing page on 2026-07-02.

View source ↗

Facts

Specs

  • FamilyQwen
  • API IDqwen-flash
  • ModalityText
  • Context window
  • Max output
  • Release dateNot published

Pricing (per 1M tokens)

  • Input$0.05 / 1M tokens
  • Output$0.40 / 1M tokens
  • Cached input
  • Cache write
  • Batch discount

Notes

International pricing, base tier (0-256K); 256K-1M tier is $0.25/$2 per 1M tokens. Alibaba recommends this as the replacement for Qwen-Turbo.

Worked cost examples

Estimated monthly cost on three common usage shapes

WorkloadEst. monthly cost

Chatbot

800 input / 200 output tokens per request, 100,000 requests/month

$12.00

RAG pipeline

6,000 input / 700 output tokens per request, 30,000 requests/month

$17.40

Batch summarization

20,000 input / 1,500 output tokens per request, 5,000 requests/month

$8.00

Price history

Changes we have documented since we began tracking

No price changes recorded since we began tracking.

Qwen-Flash pricing FAQ

How much does Qwen-Flash cost per 1M tokens?
Qwen-Flash costs $0.05 per 1M input tokens and $0.40 per 1M output tokens.
How much does Qwen-Flash cost per request?
A typical request of 800 input and 200 output tokens costs about $0.00012 on Qwen-Flash.
Is Qwen-Flash cheaper than Qwen-Long?
Qwen-Flash is cheaper than Qwen-Long on input tokens, by about 31%.
Does Qwen-Flash have cache pricing?
No cache discount is documented for Qwen-Flash on the provider's pricing page.