Baseline seeded — waiting for the first market change.
GPU market syncing — no offers yet. Admin → Bootstrap GPU market.
Head to head

Z.ai: GLM 5.3 Flash vs Xiaomi: MiMo-V2.5

Live prices, context windows, capabilities, and measured quality — updated continuously.

Side by side

MetricZ.ai: GLM 5.3 Flash
Z Ai
Xiaomi: MiMo-V2.5
Xiaomi
Input $/M tokens$0.075$0.140
Output $/M tokens$0.250$0.280
Context window1.3M1.1M
Max output131K131K
WM Score8980
Benchmark quality83.170.9
Hosts serving it245
Throughput (tok/s)8733
Latency (ms)4921175
Vision
Tools
Reasoning
JSON mode

Which should you pick?

Z.ai: GLM 5.3 Flash costs less per million input tokens — about 1.9× cheaper than Xiaomi: MiMo-V2.5. Z.ai: GLM 5.3 Flash scores higher on measured quality. Z.ai: GLM 5.3 Flash takes a larger context window, which matters for long documents or big codebases. Prices here are live list prices; the same model can cost less on a different host, so check the model page for the cheapest offering.