Expected token totals are based on typical cached agent usage and are not fixed allowances. Actual totals vary by model mix, prompt shape, output length, and cache reads. Web searches are separate tool calls and are not included in token estimates. Yearly plans are billed annually and provide the same monthly spend budget as the monthly plan. Monthly spend resets on the same calendar cadence for both monthly and yearly subscribers. Cancellation takes effect at the end of the current billing period — no prorated refunds; access remains active until the period ends.
· · ·
Tool pricing
Web search
Live web results for coding agents
Each web search is a fixed-cost tool call with zero model tokens. The charge comes from the same monthly plan or usage-pack spend balance as inference.
$0.02
per web search
· · ·
Per-model rates
Token pricing
GLM-5.2
GLM 5.2 is a long-context reasoning model for agentic coding and tool-heavy workflows.
glm-5.2·Context 524K tokens·From Z.ai
Input
$1.10
per 1M
Output
$4.00
per 1M
Cache read
$0.20
per 1M
DeepSeek V4 Flash 0731
Million-token reasoning model for coding, agents, and high-throughput workflows.