GLM 5.2 vs DeepSeek V4 Flash 0731: API Pricing Comparison
Zhipu GLM 5.2 vs DeepSeek DeepSeek V4 Flash 0731: token prices, batch and cache rates, Arena ELO, and cost per request for 8 business use cases. Prices in USD per 1 million tokens. Data snapshot: 2026-09-28.
GLM 5.2 pricing (Zhipu)
- Input: $0.65 per 1M tokens
- Output: $2.04 per 1M tokens
- Batch API (input/output): not published
- Prompt-cache read: $0.12
- Context window: 1M
- Arena ELO: 1476 (as of 2026-09-25)
- Full GLM 5.2 pricing page
DeepSeek V4 Flash 0731 pricing (DeepSeek)
- Input: $0.02 per 1M tokens
- Output: $0.32 per 1M tokens
- Batch API (input/output): not published
- Prompt-cache read: $0.02
- Context window: 1M
- Full DeepSeek V4 Flash 0731 pricing page
Cost per 1,000 requests by business use case
List price vs after discounts (prompt caching on the cached input share, batch API for async workloads, only where the provider publishes those rates).
| Use case | Tokens (in/out) | GLM 5.2 list price | GLM 5.2 after discounts | DeepSeek V4 Flash 0731 list price | DeepSeek V4 Flash 0731 after discounts |
|---|---|---|---|---|---|
| Support Ticket | 1,500 / 500 | $2.00 | $1.52 | $0.19 | $0.19 |
| Knowledge Q&A | 2,000 / 800 | $2.93 | $2.19 | $0.30 | $0.29 |
| Meeting Summary | 10,000 / 1,200 | $8.95 | $8.42 | $0.59 | $0.59 |
| Marketing Content | 2,500 / 1,800 | $5.30 | $5.03 | $0.63 | $0.63 |
| Coding Task | 3,000 / 2,000 | $6.03 | $5.24 | $0.70 | $0.70 |
| Invoice Processing | 1,500 / 600 | $2.20 | $1.96 | $0.22 | $0.22 |
| Call Summary | 2,000 / 700 | $2.73 | $2.62 | $0.27 | $0.27 |
| Agent Workflow | 6,000 / 3,000 | $10.02 | $7.80 | $1.09 | $1.07 |
Open this comparison in the interactive compare tool: add up to two more models, copy a shareable link, or export the table as CSV.
These are list prices: real cost also depends on caching, batching, negotiated discounts, marketplace agreements, and existing contracts. Get a FinOps review from OptimNow or compare all models.