Baseline seeded — waiting for the first market change.
GPU market syncing — no offers yet. Admin → Bootstrap GPU market.
Head to head

Google: Gemini 3.6 Flash vs Z.ai: GLM 5

Live prices, context windows, capabilities, and measured quality — updated continuously.

Side by side

MetricGoogle: Gemini 3.6 Flash
Google Ai Studio
Z.ai: GLM 5
Z Ai
Input $/M tokens$0.750$0.600
Output $/M tokens$3.75$1.92
Context window1.0M205K
Max output66K128K
WM Score8081
Benchmark quality80.176.7
Hosts serving it211
Throughput (tok/s)154107.5
Latency (ms)2095817
Vision
Tools
Reasoning
JSON mode

Which should you pick?

Z.ai: GLM 5 costs less per million input tokens — about 1.3× cheaper than Google: Gemini 3.6 Flash. Google: Gemini 3.6 Flash scores higher on measured quality. Google: Gemini 3.6 Flash takes a larger context window, which matters for long documents or big codebases. Prices here are live list prices; the same model can cost less on a different host, so check the model page for the cheapest offering.