GLM 5.3 Flash

Z.ai· released 26 Aug 2026

z-ai/glm-5.3-flash
Input / 1M tokens
$0.15
Output / 1M tokens
$0.5
Context window1M tokens

Longest answer: 944K tokens

reads textreads imagereads videoopen weightstool callingvisionreasoning

About

GLM 5.3 Flash is a Z.ai native multimodal model aimed at efficient coding and long-horizon agent tasks. It accepts text, images, and video, and returns text. The model supports vision, tools, and reasoning, with a context window of 1,048,576 tokens and a maximum output of 943,717 tokens. It is listed at $0.15 USD per 1M input tokens and $0.5 USD per 1M output tokens. Its model ID is z-ai/glm-5.3-flash, and its Hugging Face ID is zai-org/GLM-5.3-Flash. Open weights are listed. Z.ai lists its release and creation date as 2026-08-26. For stack decisions, the combination of video input, tool support, and a context window above one million tokens distinguishes its stated scope; output is text. The listed pricing separates input and output token rates.

Who it is for

It suits teams working on coding and long-running agent tasks that can use text, image, or video inputs. Open weights and tool support may matter to builders evaluating model integration options.

What is good

  • Accepts text, image, and video inputs
  • Open weights are listed
  • 1,048,576-token context window
  • Tool support and reasoning are listed

What to know first

  • Output is text only
  • Output costs $0.5 USD per 1M tokens

Verdict

GLM 5.3 Flash pairs multimodal inputs and open weights with a large context window. Its listed prices are $0.15 USD per 1M input tokens and $0.5 USD per 1M output tokens.

Details

Lab
Z.aiopenrouter.ai · 3 Oct 2026
Context
1,049Kopenrouter.ai · 3 Oct 2026
Input price
$0.15 / 1Mopenrouter.ai · 3 Oct 2026
Output price
$0.5 / 1Mopenrouter.ai · 3 Oct 2026
Max output
943,717 tokensopenrouter.ai · 3 Oct 2026
Inputs
text, image, videoopenrouter.ai · 3 Oct 2026
Open weights
Yesopenrouter.ai · 3 Oct 2026

More newest models

See the list