Home
Alibaba Unveils Qwen3.7-Plus, Upgrading Multimodal Agent with Enhanced Vision and Workflow

The global push for large model technology to evolve toward embodied intelligence and advanced agents is accelerating. On June 2, Alibaba announced via the Qwen Large Model official channel the official launch of Qwen3.7-Plus, a new multimodal agent model. This marks not only the next technological breakthrough for the Tongyi Qianwen series in multimodal capabilities, but also an upgrade to the core foundation for domestic large models in edge-side and complex workflow applications.
As the highlight of this upgrade, Qwen3.7-Plus builds on Qwen3.7's strong native text processing capabilities and undergoes a comprehensive, advanced evolution in vision-language abilities. This means the model can not only better interpret complex image and video content, but also convert this refined visual perception into deep logical reasoning, greatly expanding the practical boundaries of multimodal interaction.
Beyond the vision capability transformation, the model maintains top-tier strengths in the core agent chain. In areas such as programming code generation, complex tool use, and high-level productivity workflows, Qwen3.7-Plus demonstrates strong task continuity and decision robustness, enabling smoother adaptation to enterprise-level automation tasks and long-term intelligent scheduling scenarios.
Industry analysts point out that competition in the latter half of the large model era has clearly shifted toward multimodal and agent-based solutions. By deeply integrating visual understanding with agent action planning, Alibaba's Qwen3.7-Plus not only raises the performance ceiling for open-source and commercial models, but also provides a more imaginative computing foundation for broader industrial intelligence and embodied robot applications in the future.
Related article
DeepMind CEO Hassabis: I sleep six hours a day, usually feel energetic around 1 a.m.
Fortune recently featured an interview with Demis Hassabis, CEO of Google DeepMind, revealing his unconventional approach to rest and productivity. Hassabis disclosed that he sleeps very little, structuring his waking hours into two distinct work blo
OpenAI, Anthropic Vie for Market Share Despite Revenue Shortfalls
Despite recent reports suggesting OpenAI missed revenue targets, creating pressure on tech stocks this Tuesday, private AI lab investors remain resilient. Seasoned backers have confirmed they will not reduce investment despite negative media coverage
California AV Compliance: A New Era of Tickets, Geofences, and 1M Miles
Guident operates an AuveTech shuttle in South Florida, managing a four-mile route in West Palm Beach and a one-mile route in Boca Raton using its remote monitoring technology. | Credit: GuidentCalifornia is redefining the regulatory landscape for dri
Related Special Topic Recommendations
Comments (0)
0/500

The global push for large model technology to evolve toward embodied intelligence and advanced agents is accelerating. On June 2, Alibaba announced via the Qwen Large Model official channel the official launch of Qwen3.7-Plus, a new multimodal agent model. This marks not only the next technological breakthrough for the Tongyi Qianwen series in multimodal capabilities, but also an upgrade to the core foundation for domestic large models in edge-side and complex workflow applications.
As the highlight of this upgrade, Qwen3.7-Plus builds on Qwen3.7's strong native text processing capabilities and undergoes a comprehensive, advanced evolution in vision-language abilities. This means the model can not only better interpret complex image and video content, but also convert this refined visual perception into deep logical reasoning, greatly expanding the practical boundaries of multimodal interaction.
Beyond the vision capability transformation, the model maintains top-tier strengths in the core agent chain. In areas such as programming code generation, complex tool use, and high-level productivity workflows, Qwen3.7-Plus demonstrates strong task continuity and decision robustness, enabling smoother adaptation to enterprise-level automation tasks and long-term intelligent scheduling scenarios.
Industry analysts point out that competition in the latter half of the large model era has clearly shifted toward multimodal and agent-based solutions. By deeply integrating visual understanding with agent action planning, Alibaba's Qwen3.7-Plus not only raises the performance ceiling for open-source and commercial models, but also provides a more imaginative computing foundation for broader industrial intelligence and embodied robot applications in the future.
DeepMind CEO Hassabis: I sleep six hours a day, usually feel energetic around 1 a.m.
Fortune recently featured an interview with Demis Hassabis, CEO of Google DeepMind, revealing his unconventional approach to rest and productivity. Hassabis disclosed that he sleeps very little, structuring his waking hours into two distinct work blo
OpenAI, Anthropic Vie for Market Share Despite Revenue Shortfalls
Despite recent reports suggesting OpenAI missed revenue targets, creating pressure on tech stocks this Tuesday, private AI lab investors remain resilient. Seasoned backers have confirmed they will not reduce investment despite negative media coverage











