Z.AI
BigGo Finance's September 18, 2026 report independently corroborates Zhipu's launch of GLM-5.3-FlashX, describing it as the high-speed variant of the GLM-5.3 series that boosts inference speed from a prior 30–50 tokens/second to a maximum of 200 tokens/second — roughly a five-to-six-fold improvement. It states FlashX u The article adds infrastructure context that the speed gains are underpinned by an inference cluster built on more than 100,000 Chinese-made AI accelerators, with Zhipu's GLM team disclosing the cluster was built from scratch for the prior GLM generation. It also reports that InfraAgent, powered by GLM-5.3, participate