Head to head
DeepSeek: DeepSeek V4 Flash 0423 vs Z.ai: GLM 5.2
Live prices, context windows, capabilities, and measured quality — updated continuously.
Side by side
| Metric | DeepSeek: DeepSeek V4 Flash 0423 Deepseek | Z.ai: GLM 5.2 Z Ai |
|---|---|---|
| Input $/M tokens | $0.220 | $0.462 |
| Output $/M tokens | $0.660 | $1.45 |
| Context window | 1.0M | 1.0M |
| Max output | 384K | 131K |
| WM Score | 86 | 85 |
| Benchmark quality | 81 | 82.5 |
| Hosts serving it | 18 | 27 |
| Throughput (tok/s) | 82 | 113 |
| Latency (ms) | 725 | 470 |
| Vision | ||
| Tools | ||
| Reasoning | ||
| JSON mode |
Which should you pick?
DeepSeek: DeepSeek V4 Flash 0423 costs less per million input tokens — about 2.1× cheaper than Z.ai: GLM 5.2. Z.ai: GLM 5.2 scores higher on measured quality. Prices here are live list prices; the same model can cost less on a different host, so check the model page for the cheapest offering.
