ColinBuilds · Current official Z.ai API pricing
GLM-5.3
- Input / 1M
- $1.40
- Output / 1M
- $4.40
- Official context window
- 1,048,576 tokens
Current Z.ai official GLM-5.3 API pricing; input $1.40/1M, cached input $0.26/1M, output $4.40/1M. Cached-input storage listed as limited-time free. · as of August 25, 2026
Compare models
Same per-1M input/output rates as above. Enter your monthly token volumes to compare bills.
GLM-5.3
- Input: 1M × $1.40 = $1.40
- Output: 1M × $4.40 = $4.40
- Total
- $5.80
GPT-4o
- Input: 1M × $2.50 = $2.50
- Output: 1M × $10.00 = $10.00
- Total
- $12.50
Difference
$6.70
GPT-4o costs $6.70 more than GLM-5.3.
Popular comparisons
37 total featuring GLM-5.3.
GLM-5.3 is Z.ai’s current flagship text API model. Z.ai says it uses the same base model as GLM-5.2, with post-training improvements aimed at complex software engineering and agent work.
What this model is
Z.ai lists the API model ID as glm-5.3. The GLM-5.3 docs describe a 1M-token context window, 128K maximum output, text-only inputs, and reasoning that is always enabled, with reasoning_effort of low, high, or max. Disabling reasoning is not supported.
GLM-5.2 remains listed on Z.ai’s pricing page at the same token rates and is kept in this directory as the previous snapshot.
Pricing notes
Z.ai’s pricing page lists GLM-5.3 at $1.40 per 1M input tokens, $0.26 per 1M cached input tokens, limited-time free cached-input storage, and $4.40 per 1M output tokens.
The calculator on this page uses the standard input and output prices: $1.40 input and $4.40 output per 1M tokens. Cached-input pricing is listed separately because it depends on workload shape.
Mistral’s API catalog also hosts GLM 5.2 at the same $1.40 / $4.40 token rates under zai-glm-5-2. This page uses first-party Z.ai pricing for GLM-5.3.
Benchmarks and specs
Z.ai’s GLM-5.3 documentation lists a 1M context length and 128K maximum output tokens, text input and text output, streaming, function calling, and always-on reasoning.
This page does not list specific benchmark scores until they are tied to a named source and test setup. Z.ai’s own coding-benchmark claims are therefore omitted here.
Best fit
GLM-5.3 is best for current Z.ai flagship comparisons on long-horizon coding and agent tasks where official Z.ai API pricing is required.
For the previous snapshot at the same token price, compare against GLM-5.2.