Head to head
Google: Gemini 3.5 Flash vs Z.ai: GLM 5
Live prices, context windows, capabilities, and measured quality — updated continuously.
Side by side
| Metric | Google: Gemini 3.5 Flash Google Ai Studio | Z.ai: GLM 5 Z Ai |
|---|---|---|
| Input $/M tokens | $1.50 | $1.00 |
| Output $/M tokens | $9.00 | $3.20 |
| Context window | 1.0M | 205K |
| Max output | 66K | 128K |
| WM Score | 81 | 81 |
| Benchmark quality | 79.6 | 76.6 |
| Hosts serving it | 2 | 11 |
| Throughput (tok/s) | 56 | 106.5 |
| Latency (ms) | 798 | 747 |
| Vision | ||
| Tools | ||
| Reasoning | ||
| JSON mode |
Which should you pick?
Z.ai: GLM 5 costs less per million input tokens — about 1.5× cheaper than Google: Gemini 3.5 Flash. Google: Gemini 3.5 Flash scores higher on measured quality. Google: Gemini 3.5 Flash takes a larger context window, which matters for long documents or big codebases. Prices here are live list prices; the same model can cost less on a different host, so check the model page for the cheapest offering.
