GLM 5.2 vs DeepSeek V4 Flash 0731: API Pricing Comparison

Zhipu GLM 5.2 vs DeepSeek DeepSeek V4 Flash 0731 — token prices, batch and cache rates, Arena ELO, and cost per request for 8 business use cases. Prices in USD per 1 million tokens. Data snapshot: 2026-08-14.

GLM 5.2 pricing (Zhipu)

DeepSeek V4 Flash 0731 pricing (DeepSeek)

Cost per request by business use case

List price vs optimized (prompt caching on the cached input share, batch API for async workloads — only where the provider publishes those rates).

Use caseTokens (in/out)GLM 5.2 listGLM 5.2 optimizedDeepSeek V4 Flash 0731 listDeepSeek V4 Flash 0731 optimized
Support Ticket1,500 / 500$0.0019$0.0015$0.000350$0.000249
Knowledge Q&A2,000 / 800$0.0028$0.0021$0.000504$0.000347
Meeting Summary10,000 / 1,200$0.0087$0.0090$0.0017$0.0016
Marketing Content2,500 / 1,800$0.0051$0.0049$0.000854$0.000798
Coding Task3,000 / 2,000$0.0059$0.0050$0.000980$0.000812
Invoice Processing1,500 / 600$0.0021$0.0021$0.000378$0.000328
Call Summary2,000 / 700$0.0026$0.0028$0.000476$0.000454
Agent Workflow6,000 / 3,000$0.0097$0.0075$0.0017$0.0012

Open this comparison in the interactive compare tool — add up to two more models, copy a shareable link, or export the table as CSV.

These are list prices — real cost also depends on caching, batching, negotiated discounts, marketplace agreements, and existing contracts. Get a FinOps review from OptimNow or compare all models.