Zro

GLM-5.3 Flash

Fast, million-token reasoning model for coding agents and tool-heavy workflows.

Fast, million-token reasoning model for coding agents and tool-heavy workflows.

Model IDglm-5.3-flash
Context window1M tokens
Max output64K tokens
Modalitiestext, image
Aliaszro/glm-5.3-flash

Pricing

GLM-5.3 Flash pricing: $0.15 per 1M input tokens, $0.50 per 1M output tokens, and $0.03 per 1M cache read tokens.

InputOutputCache read
$0.15$0.50$0.03

Reasoning effort

Default: max. Available levels: none, high, max.

Upstream weights

zai-org/GLM-5.3-Flash (MIT License).

On this page