Grok 4.20 Multi Agent API pricing
xAI model · context 1,000,000 tokens · max output 1,000,000 · vision, reasoning. Data last changed 2026-10-08. Data synced from the LiteLLM price list 2026-10-08.
| Input | $1.25 / 1M tokens |
|---|---|
| Output | $2.50 / 1M tokens |
| Cached input | $0.20 / 1M tokens |
| Batch input | $1.00 / 1M tokens |
| Batch output | $2.00 / 1M tokens |
| Batch cached input | $0.16 / 1M tokens |
| Input above 200K | $2.50 / 1M tokens |
| Output above 200K | $5.00 / 1M tokens |
| Long-context cached input (>200K) | $0.40 / 1M tokens |
| Long-context batch input (>200K) | $2.00 / 1M tokens |
| Long-context batch output (>200K) | $4.00 / 1M tokens |
| Long-context batch cached input (>200K) | $0.32 / 1M tokens |
What it costs per month
| Scenario (per month) | Grok 4.20 Multi Agent |
|---|---|
| Customer support chatbot 5,000 req/day · 1,200 in · 400 out · cache 60% | $262 |
| RAG over documents 2,000 req/day · 8,000 in · 600 out · cache 30% | $539 |
| Coding agent 1,500 req/day · 25,000 in · 2,000 out · cache 80% | $686 |
| Deep reasoning 800 req/day · 3,000 in · 6,000 out · cache 20% | $435 |
| Batch processing 50,000 req/day · 1,500 in · 150 out · batch | $2,700 |
Adjust the numbers in the calculator.
Cheaper alternatives
Lowest monthly cost for the chatbot scenario among reasoning models:
- Qwen Turbo — $21.00/mo
- MiMo V2.6 Flash — $27.18/mo
- GPT-5 nano — $28.14/mo
- Gemini 2.5 Flash Lite — $32.28/mo
- Qwen Flash — $33.00/mo
Comparisons
Price history
- 2026-10-08: Added — → —