Z.ai Models: Pricing per 1M Tokens and Context (2026)

Z.ai’s catalogue brings together its models for readers comparing token pricing and context across model variants. Browse the 20 entries to identify options relevant to your workload, and use the available filters and sorting to narrow the catalogue. The entries are ordered by creation date, newest first, so you can see newer releases before older ones. GLM 5.3 Prime, GLM 5.3 FlashX, and GLM Flash Latest are among the models included. Compare their listed pricing per million tokens and context details against the requirements of your application.

20 on record. Sorted by created; every figure comes from the OpenRouter model catalogue, with the page it was read from on each entry.

20listed

Context bars share one log scale, 1K to 2M tokens. Prices are what each provider publishes per million tokens; 'varies' is a router that bills the model it picks.

More in By lab

All by lab lists