GLM 5.2 vs DeepSeek V4 Flash 0731: API Pricing Comparison
Zhipu GLM 5.2 vs DeepSeek DeepSeek V4 Flash 0731 — token prices, batch and cache rates, Arena ELO, and cost per request for 8 business use cases. Prices in USD per 1 million tokens. Data snapshot: 2026-08-14.
GLM 5.2 pricing (Zhipu)
- Input: $0.63 per 1M tokens
- Output: $1.98 per 1M tokens
- Batch API (input/output): $0.70 / $2.20
- Prompt-cache read: $0.10
- Context window: 1M
- Arena ELO: 1430 (as of 2026-08-13)
- Full GLM 5.2 pricing page
DeepSeek V4 Flash 0731 pricing (DeepSeek)
- Input: $0.14 per 1M tokens
- Output: $0.28 per 1M tokens
- Batch API (input/output): not published
- Prompt-cache read: $0.03
- Context window: 1M
- Full DeepSeek V4 Flash 0731 pricing page
Cost per request by business use case
List price vs optimized (prompt caching on the cached input share, batch API for async workloads — only where the provider publishes those rates).
| Use case | Tokens (in/out) | GLM 5.2 list | GLM 5.2 optimized | DeepSeek V4 Flash 0731 list | DeepSeek V4 Flash 0731 optimized |
|---|---|---|---|---|---|
| Support Ticket | 1,500 / 500 | $0.0019 | $0.0015 | $0.000350 | $0.000249 |
| Knowledge Q&A | 2,000 / 800 | $0.0028 | $0.0021 | $0.000504 | $0.000347 |
| Meeting Summary | 10,000 / 1,200 | $0.0087 | $0.0090 | $0.0017 | $0.0016 |
| Marketing Content | 2,500 / 1,800 | $0.0051 | $0.0049 | $0.000854 | $0.000798 |
| Coding Task | 3,000 / 2,000 | $0.0059 | $0.0050 | $0.000980 | $0.000812 |
| Invoice Processing | 1,500 / 600 | $0.0021 | $0.0021 | $0.000378 | $0.000328 |
| Call Summary | 2,000 / 700 | $0.0026 | $0.0028 | $0.000476 | $0.000454 |
| Agent Workflow | 6,000 / 3,000 | $0.0097 | $0.0075 | $0.0017 | $0.0012 |
Open this comparison in the interactive compare tool — add up to two more models, copy a shareable link, or export the table as CSV.
These are list prices — real cost also depends on caching, batching, negotiated discounts, marketplace agreements, and existing contracts. Get a FinOps review from OptimNow or compare all models.