Qwen3.7 Plus API pricing

Alibaba (Qwen) model · context 991,808 tokens · max output 65,536 · vision, reasoning. Data last changed 2026-10-08. Official pricing ↗ · Data synced from the LiteLLM price list 2026-10-08.

Input$0.40 / 1M tokens
Output$1.60 / 1M tokens
Cached input$0.08 / 1M tokens
Input above 256K$1.20 / 1M tokens
Output above 256K$4.80 / 1M tokens
Long-context cached input (>256K)$0.24 / 1M tokens

Calculate my cost →

What it costs per month

Scenario (per month)Qwen3.7 Plus
Customer support chatbot
5,000 req/day · 1,200 in · 400 out · cache 60%
$133
RAG over documents
2,000 req/day · 8,000 in · 600 out · cache 30%
$204
Coding agent
1,500 req/day · 25,000 in · 2,000 out · cache 80%
$306
Deep reasoning
800 req/day · 3,000 in · 6,000 out · cache 20%
$255
Batch processing
50,000 req/day · 1,500 in · 150 out · batch
$1,260

Adjust the numbers in the calculator.

Cheaper alternatives

Lowest monthly cost for the chatbot scenario among reasoning models:

Comparisons

Price history