GLM-5 API pricing

Z.ai model · context 200,000 tokens · max output 128,000 · reasoning. Data last changed 2026-10-08. Official pricing ↗ · Data synced from the LiteLLM price list 2026-10-08.

Input$1.00 / 1M tokens
Output$3.20 / 1M tokens
Cached input$0.20 / 1M tokens
Cache write$0.00 / 1M tokens

Calculate my cost →

What it costs per month

Scenario (per month)GLM-5
Customer support chatbot
5,000 req/day · 1,200 in · 400 out · cache 60%
$277
RAG over documents
2,000 req/day · 8,000 in · 600 out · cache 30%
$456
Coding agent
1,500 req/day · 25,000 in · 2,000 out · cache 80%
$580
Deep reasoning
800 req/day · 3,000 in · 6,000 out · cache 20%
$521
Batch processing
50,000 req/day · 1,500 in · 150 out · batch
$2,970

Adjust the numbers in the calculator.

Cheaper alternatives

Lowest monthly cost for the chatbot scenario among reasoning models:

Comparisons

Price history