Best Input Price
$0.37 / 1M
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
Context Window
1,048,576 tokens
Reasoning
Supported
Tool Calling
Supported
Released
2026-09-18
Inference Providers (2)
| Provider | Model ID | Context | Input / 1M | Output / 1M | Action |
|---|---|---|---|---|---|
| Kilo Gateway | z-ai/glm-5.3-flashx | 1,048,576 | $0.37 | $1.25 | Docs ↗ |
| OpenRouter | z-ai/glm-5.3-flashx | 1,048,576 | $0.37 | $1.25 | Docs ↗ |