OpenRouter
Superpower Daily reported on September 18, 2026 that Z.AI added GLM-5.3-Flash to its Coding Plan with three times the GLM-5.3 quota, but deliberately excluded the faster FlashX variant from the subscription for now. Both models are available through Z.AI's API under the identifiers glm-5.3-flash and glm-5.3-flashx resp The same article describes GLM-5.3-Flash as the first native multimodal model in the GLM-5 line, accepting video, images, text, and files with a one-million-token context window, 320B total / 18B active parameters, and substantially lower attention and KV-cache costs on long prompts compared with GLM-5.3. FlashX's adve