Baseline seeded — waiting for the first market change.
GPU market syncing — no offers yet. Admin → Bootstrap GPU market.
Back to explorer
Z Ai

Z.ai: GLM 5.3 Flash

Z Ai

Input / M
Output / M
Context
Max Output
131K
Usage
Unranked
Quality
Unscored
Providers
1
Best Uptime 30m
100.00%
Latency (p50)
5173 ms
Throughput
33.0 tok/s
Released
August 26, 2026
Modalities
text, image, video

About

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while.

Capabilities

Vision
Tools
Reasoning
Streaming
JSON Mode

Host offerings

Per-host pricing, uptime, latency
No host data yet — will populate on next sync.

Price History

0 recorded changes
No price changes recorded yet.

Activity Timeline

  • No events yet.

Versus

Nearest peers by price, context, and capabilities. One click to compare.
No comparable models found.

Show this WM Score

Free live badge for your README, docs or site
Z.ai: GLM 5.3 Flash WM Score badge

Live badge — it always shows the current WM Score. Free to use anywhere, no key needed.

Markdown
[![Z.ai: GLM 5.3 Flash WM Score](https://whatmodel.app/api/public/embed/score/42fb9801-a210-4ac3-b727-ca7138735d37)](https://whatmodel.app/models/42fb9801-a210-4ac3-b727-ca7138735d37)
HTML
<a href="https://whatmodel.app/models/42fb9801-a210-4ac3-b727-ca7138735d37"><img src="https://whatmodel.app/api/public/embed/score/42fb9801-a210-4ac3-b727-ca7138735d37" alt="Z.ai: GLM 5.3 Flash WM Score" height="26" /></a>
Image URL
https://whatmodel.app/api/public/embed/score/42fb9801-a210-4ac3-b727-ca7138735d37
First seen 1h ago · Last verified 9m ago