Baseline seeded — waiting for the first market change.
GPU market syncing — no offers yet. Admin → Bootstrap GPU market.
Head to head

Z.ai: GLM 5.2 vs Google: Gemini 3.5 Flash

Live prices, context windows, capabilities, and measured quality — updated continuously.

Side by side

MetricZ.ai: GLM 5.2
Z Ai
Google: Gemini 3.5 Flash
Google
Input $/M tokens$1.40$1.50
Output $/M tokens$4.40$9.00
Context window1.0M1.0M
Max output131K66K
WM Score8581
Benchmark quality82.579.6
Hosts serving it272
Throughput (tok/s)19689.5
Latency (ms)327909
Vision
Tools
Reasoning
JSON mode

Which should you pick?

Z.ai: GLM 5.2 costs less per million input tokens — about 1.1× cheaper than Google: Gemini 3.5 Flash. Z.ai: GLM 5.2 scores higher on measured quality. Prices here are live list prices; the same model can cost less on a different host, so check the model page for the cheapest offering.