IteraCompute
Z.ai officially announced GLM-5.3-Flash on August 26, 2026, describing it as the first natively multimodal model in the GLM-5 series. The model is a mixture-of-experts with 320B total parameters and 18B active per token, offering a 1,048,576-token context window and supporting image and video input. Z.ai states GLM-5.3 On the Artificial Analysis Intelligence Index v4.1.1, Z.ai reports GLM-5.3-Flash scores 57 at a discounted cost of $0.045 per task. On coding benchmarks it posts 63.4 on DeepSWE v1.1 (up from 46.2 for GLM-5.2) and 48.8 on AutomationBench (up from 26.2), and on Z.ai Code Bench v1.0 it scores 29.0 at max effort versus 29