Head to head
Z.ai: GLM 5.2 vs Google: Gemini 3.7 Flash
Live prices, context windows, capabilities, and measured quality — updated continuously.
Side by side
| Metric | Z.ai: GLM 5.2 Z Ai | Google: Gemini 3.7 Flash Google Ai Studio |
|---|---|---|
| Input $/M tokens | $1.40 | $0.750 |
| Output $/M tokens | $4.40 | $3.75 |
| Context window | 1.0M | 1.0M |
| Max output | 262K | 66K |
| WM Score | 85 | 85 |
| Benchmark quality | 82.4 | 86.9 |
| Hosts serving it | 26 | 2 |
| Throughput (tok/s) | 138 | 181 |
| Latency (ms) | 378 | 3193 |
| Vision | ||
| Tools | ||
| Reasoning | ||
| JSON mode |
Which should you pick?
Google: Gemini 3.7 Flash costs less per million input tokens — about 1.9× cheaper than Z.ai: GLM 5.2. Google: Gemini 3.7 Flash scores higher on measured quality. Prices here are live list prices; the same model can cost less on a different host, so check the model page for the cheapest offering.
