GLM-5.3 Flash
Fast, million-token reasoning model for coding agents and tool-heavy workflows.
Fast, million-token reasoning model for coding agents and tool-heavy workflows.
| Model ID | glm-5.3-flash |
| Context window | 1M tokens |
| Max output | 64K tokens |
| Modalities | text, image |
| Alias | zro/glm-5.3-flash |
Pricing
GLM-5.3 Flash pricing: $0.15 per 1M input tokens, $0.50 per 1M output tokens, and $0.03 per 1M cache read tokens.
| Input | Output | Cache read |
|---|---|---|
$0.15 | $0.50 | $0.03 |
Reasoning effort
Default: max. Available levels: none, high, max.
Upstream weights
zai-org/GLM-5.3-Flash (MIT License).