DeepSeek: DeepSeek V4 Flash 0731 vs Z.ai: GLM 5.2 (batch)
Based on weekly OpenRouter usage tracked by AI Tier List, DeepSeek: DeepSeek V4 Flash 0731 ranks #1 overall — ahead of Z.ai: GLM 5.2 (batch) at #6. DeepSeek: DeepSeek V4 Flash 0731 is cheaper on input tokens at $0.14/1M vs $0.70/1M for Z.ai: GLM 5.2 (batch). DeepSeek: DeepSeek V4 Flash 0731 offers the larger context window at 1M tokens.
DeepSeek: DeepSeek V4 Flash 0731 vs Z.ai: GLM 5.2 (batch): which should you use?
Choose DeepSeek: DeepSeek V4 Flash 0731 if…
- It sees more real traffic — #1 overall on OpenRouter's weekly ranking, against #6 for Z.ai: GLM 5.2 (batch).
- Cheaper to read — $0.14 vs $0.70 per 1M input tokens, which is what dominates long-prompt and RAG workloads.
- Cheaper to write — $0.28 vs $2.20 per 1M output tokens, the line item that decides generation-heavy costs.
- Bigger context window — 1M vs 512K tokens.
- It is the newer model — released Jul 2026 against Jun 2026.
Choose Z.ai: GLM 5.2 (batch) if…
On every published figure — price, context, benchmarks — it does not come out ahead of DeepSeek: DeepSeek V4 Flash 0731.
Spec Comparison
| DeepSeek: DeepSeek V4 Flash 0731 | Z.ai: GLM 5.2 (batch) | |
|---|---|---|
| Author | deepseek | z-ai |
| Input price (per 1M tokens) | $0.14 | $0.70 |
| Output price (per 1M tokens) | $0.28 | $2.20 |
| Context window | 1M | 512K |
| Input modalities | text | text |
| Output modalities | text | text |
| AA Intelligence | 51.8 | — |
| AA Coding | 69.1 | — |
| AA Agentic | 48.4 | — |
| Released | Jul 2026 | Jun 2026 |
What DeepSeek: DeepSeek V4 Flash 0731 and Z.ai: GLM 5.2 (batch) actually cost to run
On a read-heavy job — 1M tokens in, 100K out — DeepSeek: DeepSeek V4 Flash 0731 costs $0.17 and Z.ai: GLM 5.2 (batch) costs $0.92.
Flip the shape — 100K in, 1M out — and it becomes $0.29 for DeepSeek: DeepSeek V4 Flash 0731 versus $2.27 for Z.ai: GLM 5.2 (batch).
Both figures are the published per-token rates multiplied out; caching and batch discounts are not included.
Usage Rank by Task
| Task | DeepSeek: DeepSeek V4 Flash 0731 | Z.ai: GLM 5.2 (batch) |
|---|---|---|
| Overall | #1 | #6 |
| Memory Extraction | #6 | #7 |
| Multi-step Planning | #2 | #5 |
| Tool Dispatch | #2 | #5 |
| Web Search | #4 | #6 |
| Agent Workflows | #5 | #4 |
| Classification & Tagging | #9 | — |
| Debugging | #3 | #2 |
| DevOps Config | #2 | #3 |
| File Read/Write | #3 | #5 |
| Frontend UI | #2 | #3 |
| Code Implementation | #3 | #2 |
| Repo Analysis | #3 | #7 |
| Code Review & Security | #3 | #1 |
| Shell Execution | #6 | #5 |
About DeepSeek: DeepSeek V4 Flash 0731 and Z.ai: GLM 5.2 (batch)
DeepSeek: DeepSeek V4 Flash 0731
Built by deepseek · released Jul 2026 · 1M-token context
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
Z.ai: GLM 5.2 (batch)
Built by z-ai · released Jun 2026 · 512K-token context
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
Get notified when DeepSeek: DeepSeek V4 Flash 0731 or Z.ai: GLM 5.2 (batch) moves
Usage ranks and prices refresh weekly. We email you when they move — once a week, no spam.
FAQ
Which is better, DeepSeek: DeepSeek V4 Flash 0731 or Z.ai: GLM 5.2 (batch)?
Based on weekly OpenRouter usage tracked by AI Tier List, DeepSeek: DeepSeek V4 Flash 0731 ranks #1 overall — ahead of Z.ai: GLM 5.2 (batch) at #6. DeepSeek: DeepSeek V4 Flash 0731 is cheaper on input tokens at $0.14/1M vs $0.70/1M for Z.ai: GLM 5.2 (batch). DeepSeek: DeepSeek V4 Flash 0731 offers the larger context window at 1M tokens.
Which is cheaper, DeepSeek: DeepSeek V4 Flash 0731 or Z.ai: GLM 5.2 (batch)?
DeepSeek: DeepSeek V4 Flash 0731 is cheaper at $0.14 per 1M input tokens. Output pricing: DeepSeek: DeepSeek V4 Flash 0731 $0.28, Z.ai: GLM 5.2 (batch) $2.20.
Usage ranks are weekly data aggregated by OpenRouter. Prices are USD per 1M tokens.