GLM-5.3-Flash is Zhipu AI’s natively multimodal MoE model in the GLM-5 family. With 320B total parameters and 18B active parameters, it adopts hybrid sparse and linear attention architecture and supports a 1M-token context window. It accepts text, image, video and file inputs, delivering strong coding and Agent capabilities. Open-sourced under MIT license with greatly reduced cost, it fits multimodal document parsing, UI analysis, professional content generation and agent workflows.