GLM 5.2 vs DeepSeek V4 Flash 0731: API Pricing Comparison

Zhipu GLM 5.2 vs DeepSeek DeepSeek V4 Flash 0731: token prices, batch and cache rates, Arena ELO, and cost per request for 8 business use cases. Prices in USD per 1 million tokens. Data snapshot: 2026-09-06.

GLM 5.2 pricing (Zhipu)

DeepSeek V4 Flash 0731 pricing (DeepSeek)

Cost per request by business use case

List price vs optimized (prompt caching on the cached input share, batch API for async workloads, only where the provider publishes those rates).

Use caseTokens (in/out)GLM 5.2 listGLM 5.2 optimizedDeepSeek V4 Flash 0731 listDeepSeek V4 Flash 0731 optimized
Support Ticket1,500 / 500$0.0030$0.0023$0.000125$0.000089
Knowledge Q&A2,000 / 800$0.0044$0.0033$0.000180$0.000124
Meeting Summary10,000 / 1,200$0.013$0.013$0.000620$0.000580
Marketing Content2,500 / 1,800$0.0079$0.0075$0.000305$0.000285
Coding Task3,000 / 2,000$0.0090$0.0078$0.000350$0.000290
Invoice Processing1,500 / 600$0.0033$0.0029$0.000135$0.000117
Call Summary2,000 / 700$0.0041$0.0039$0.000170$0.000162
Agent Workflow6,000 / 3,000$0.015$0.012$0.000600$0.000432

Open this comparison in the interactive compare tool: add up to two more models, copy a shareable link, or export the table as CSV.

These are list prices: real cost also depends on caching, batching, negotiated discounts, marketplace agreements, and existing contracts. Get a FinOps review from OptimNow or compare all models.