ColinBuilds

ColinBuilds · Current official Z.ai API pricing

GLM-5.3

Input / 1M
$1.40
Output / 1M
$4.40
Official context window
1,048,576 tokens

Current Z.ai official GLM-5.3 API pricing; input $1.40/1M, cached input $0.26/1M, output $4.40/1M. Cached-input storage listed as limited-time free. · as of August 25, 2026

Compare models

Same per-1M input/output rates as above. Enter your monthly token volumes to compare bills.

GLM-5.3

Input: 1M × $1.40 = $1.40
Output: 1M × $4.40 = $4.40
Total
$5.80

GPT-4o

Input: 1M × $2.50 = $2.50
Output: 1M × $10.00 = $10.00
Total
$12.50

Difference

$6.70

GPT-4o costs $6.70 more than GLM-5.3.

Popular comparisons

37 total featuring GLM-5.3.

Open compare hub →

GLM-5.3 is Z.ai’s current flagship text API model. Z.ai says it uses the same base model as GLM-5.2, with post-training improvements aimed at complex software engineering and agent work.

What this model is

Z.ai lists the API model ID as glm-5.3. The GLM-5.3 docs describe a 1M-token context window, 128K maximum output, text-only inputs, and reasoning that is always enabled, with reasoning_effort of low, high, or max. Disabling reasoning is not supported.

GLM-5.2 remains listed on Z.ai’s pricing page at the same token rates and is kept in this directory as the previous snapshot.

Pricing notes

Z.ai’s pricing page lists GLM-5.3 at $1.40 per 1M input tokens, $0.26 per 1M cached input tokens, limited-time free cached-input storage, and $4.40 per 1M output tokens.

The calculator on this page uses the standard input and output prices: $1.40 input and $4.40 output per 1M tokens. Cached-input pricing is listed separately because it depends on workload shape.

Mistral’s API catalog also hosts GLM 5.2 at the same $1.40 / $4.40 token rates under zai-glm-5-2. This page uses first-party Z.ai pricing for GLM-5.3.

Benchmarks and specs

Z.ai’s GLM-5.3 documentation lists a 1M context length and 128K maximum output tokens, text input and text output, streaming, function calling, and always-on reasoning.

This page does not list specific benchmark scores until they are tied to a named source and test setup. Z.ai’s own coding-benchmark claims are therefore omitted here.

Best fit

GLM-5.3 is best for current Z.ai flagship comparisons on long-horizon coding and agent tasks where official Z.ai API pricing is required.

For the previous snapshot at the same token price, compare against GLM-5.2.

Sources