GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
Share of successful requests per carrier, measured over the last 5 minutes, 30 minutes and day. When one carrier dips, the router fails over to the next.
| Provider | Status | Last 5m | Last 30m | Last day |
|---|---|---|---|---|
| Operational | 100.00% | 100.00% | 100.00% |