GLM 5.2 vs DeepSeek V4 Flash 0731: API Pricing Comparison
Zhipu GLM 5.2 vs DeepSeek DeepSeek V4 Flash 0731: token prices, batch and cache rates, Arena ELO, and cost per request for 8 business use cases. Prices in USD per 1 million tokens. Data snapshot: 2026-09-06.
GLM 5.2 pricing (Zhipu)
- Input: $0.97 per 1M tokens
- Output: $3.04 per 1M tokens
- Batch API (input/output): not published
- Prompt-cache read: $0.19
- Context window: 1M
- Arena ELO: 1472 (as of 2026-09-01)
- Full GLM 5.2 pricing page
DeepSeek V4 Flash 0731 pricing (DeepSeek)
- Input: $0.05 per 1M tokens
- Output: $0.10 per 1M tokens
- Batch API (input/output): not published
- Prompt-cache read: $0.01
- Context window: 1M
- Full DeepSeek V4 Flash 0731 pricing page
Cost per request by business use case
List price vs optimized (prompt caching on the cached input share, batch API for async workloads, only where the provider publishes those rates).
| Use case | Tokens (in/out) | GLM 5.2 list | GLM 5.2 optimized | DeepSeek V4 Flash 0731 list | DeepSeek V4 Flash 0731 optimized |
|---|---|---|---|---|---|
| Support Ticket | 1,500 / 500 | $0.0030 | $0.0023 | $0.000125 | $0.000089 |
| Knowledge Q&A | 2,000 / 800 | $0.0044 | $0.0033 | $0.000180 | $0.000124 |
| Meeting Summary | 10,000 / 1,200 | $0.013 | $0.013 | $0.000620 | $0.000580 |
| Marketing Content | 2,500 / 1,800 | $0.0079 | $0.0075 | $0.000305 | $0.000285 |
| Coding Task | 3,000 / 2,000 | $0.0090 | $0.0078 | $0.000350 | $0.000290 |
| Invoice Processing | 1,500 / 600 | $0.0033 | $0.0029 | $0.000135 | $0.000117 |
| Call Summary | 2,000 / 700 | $0.0041 | $0.0039 | $0.000170 | $0.000162 |
| Agent Workflow | 6,000 / 3,000 | $0.015 | $0.012 | $0.000600 | $0.000432 |
Open this comparison in the interactive compare tool: add up to two more models, copy a shareable link, or export the table as CSV.
These are list prices: real cost also depends on caching, batching, negotiated discounts, marketplace agreements, and existing contracts. Get a FinOps review from OptimNow or compare all models.