qwen3.7-plus is Alibaba Cloud’s commercial multimodal MoE model of Qwen3.7 family. With 397B total parameters and 17B active parameters, it supports a 1M-token context window. Built with native multimodal Agent capability, it processes text, image and video inputs. It can read screens, analyze GUI interfaces and generate code from visual references. Supporting thinking/non-thinking hybrid reasoning, function calling, structured output and web search, it fits multimodal document parsing, GUI automation, visual coding and cross-modal long-chain Agent workflows.