Z Ai: Z.ai: GLM 5.3 FlashX
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture... Up to 1M-token context window for long documents and extended agent traces Native text, image, and video input with text output Reasoning-enabled inference with configurable effort levels FlashX is a speed-focused variant with less headroom than larger GLM models OpenRouter pricing and availability may vary by provider Fast multimodal agents with long context Video and image understanding High-volume coding and reasoning workloads
Specifications
| Lab | Z Ai |
|---|---|
| Context window | 1,048,576 |
| Input price | $0.37/1M |
| Output price | $1.25/1M |
| Release | 2026-09-01 00:00:00 |
Strengths
- Up to 1M-token context window for long documents and extended agent traces
- Native text, image, and video input with text output
- Reasoning-enabled inference with configurable effort levels
Weaknesses
- FlashX is a speed-focused variant with less headroom than larger GLM models
- OpenRouter pricing and availability may vary by provider
Best for
- Fast multimodal agents with long context
- Video and image understanding
- High-volume coding and reasoning workloads
In Depth: Z.ai: GLM 5.3 FlashX
Summary
Z.ai: GLM 5.3 FlashX is an AI model from Z Ai.
Released 2026-09-01 00:00:00. It supports Text, image, video input and produces Text output, with a context window of 1M tokens. Input pricing is $0.37 per 1M tokens and output is $1.25 per 1M tokens on OpenRouter.