Qwen Flash API pricing

Alibaba (Qwen) model · context 997,952 tokens · max output 32,768 · reasoning. Data last changed 2026-10-08. Official pricing ↗ · Data synced from the LiteLLM price list 2026-10-08.

Input$0.05 / 1M tokens
Output$0.40 / 1M tokens
Input above 256K$0.25 / 1M tokens
Output above 256K$2.00 / 1M tokens

Calculate my cost →

What it costs per month

Scenario (per month)Qwen Flash
Customer support chatbot
5,000 req/day · 1,200 in · 400 out · cache 60%
$33.00
RAG over documents
2,000 req/day · 8,000 in · 600 out · cache 30%
$38.40
Coding agent
1,500 req/day · 25,000 in · 2,000 out · cache 80%
$92.25
Deep reasoning
800 req/day · 3,000 in · 6,000 out · cache 20%
$61.20
Batch processing
50,000 req/day · 1,500 in · 150 out · batch
$202

Adjust the numbers in the calculator.

Cheaper alternatives

Lowest monthly cost for the chatbot scenario among reasoning models:

Comparisons

Price history