Alibaba Cloud Unveils Qianwen Cloud, Enabling AI Agents to Automatically Migrate
On May 20, 2026, at the Alibaba Cloud Summit, Alibaba Cloud unveiled Qwen Cloud, a new AI service platform built for the Agentic era. Described as "the full-stack intelligent infrastructure for AI Agents," this platform signals a shift from compute-centric to agent-centric cloud computing, redefining the service pipeline in the age of large language models.

Qwen Cloud’s core highlights lie in achieving full "Skillization" and "CLIization" of model services. By packaging complex steps like model selection, resource invocation, authentication, and usage tracking into standardized tool interfaces, AI Agents can access platform capabilities with simple commands—no manual coding needed. The platform now hosts over 150 model families, including Tongyi Qianwen Qwen3.7-Max, Zhipu GLM, Moonshot Kimi, DeepSeek, and more, totaling over 480 mainstream models. Among them, the flagship Qwen3.7-Max ranks first among domestic models in blind test overall rankings, showing exceptional alignment with task objectives.
On the infrastructure side, Alibaba Cloud also launched the Panjiu server, powered by its new in-house AI chip, the Zhenwu M890. This server delivers three times the performance of its predecessor, with point-to-point latency under 150 nanoseconds. Together with the upgraded "Agentic Cloud" infrastructure, cloud products now offer lightweight sandbox environments tailored for Agents, capable of handling workloads like high-frequency, short-lived, and burst concurrency scenarios.
Additionally, the platform introduced an innovative "Token Plan" subscription model to lower costs for frequent AI programming and agent tool usage through more flexible billing. This move indicates that cloud services are entering a fully "agent-native" phase, accelerating the shift from single-generation AI applications to complex task automation.
Related article
DeepMind CEO Hassabis: I sleep six hours a day, usually feel energetic around 1 a.m.
Fortune recently featured an interview with Demis Hassabis, CEO of Google DeepMind, revealing his unconventional approach to rest and productivity. Hassabis disclosed that he sleeps very little, structuring his waking hours into two distinct work blo
OpenAI, Anthropic Vie for Market Share Despite Revenue Shortfalls
Despite recent reports suggesting OpenAI missed revenue targets, creating pressure on tech stocks this Tuesday, private AI lab investors remain resilient. Seasoned backers have confirmed they will not reduce investment despite negative media coverage
California AV Compliance: A New Era of Tickets, Geofences, and 1M Miles
Guident operates an AuveTech shuttle in South Florida, managing a four-mile route in West Palm Beach and a one-mile route in Boca Raton using its remote monitoring technology. | Credit: GuidentCalifornia is redefining the regulatory landscape for dri
Related Special Topic Recommendations
Comments (1)
0/500
On May 20, 2026, at the Alibaba Cloud Summit, Alibaba Cloud unveiled Qwen Cloud, a new AI service platform built for the Agentic era. Described as "the full-stack intelligent infrastructure for AI Agents," this platform signals a shift from compute-centric to agent-centric cloud computing, redefining the service pipeline in the age of large language models.

Qwen Cloud’s core highlights lie in achieving full "Skillization" and "CLIization" of model services. By packaging complex steps like model selection, resource invocation, authentication, and usage tracking into standardized tool interfaces, AI Agents can access platform capabilities with simple commands—no manual coding needed. The platform now hosts over 150 model families, including Tongyi Qianwen Qwen3.7-Max, Zhipu GLM, Moonshot Kimi, DeepSeek, and more, totaling over 480 mainstream models. Among them, the flagship Qwen3.7-Max ranks first among domestic models in blind test overall rankings, showing exceptional alignment with task objectives.
On the infrastructure side, Alibaba Cloud also launched the Panjiu server, powered by its new in-house AI chip, the Zhenwu M890. This server delivers three times the performance of its predecessor, with point-to-point latency under 150 nanoseconds. Together with the upgraded "Agentic Cloud" infrastructure, cloud products now offer lightweight sandbox environments tailored for Agents, capable of handling workloads like high-frequency, short-lived, and burst concurrency scenarios.
Additionally, the platform introduced an innovative "Token Plan" subscription model to lower costs for frequent AI programming and agent tool usage through more flexible billing. This move indicates that cloud services are entering a fully "agent-native" phase, accelerating the shift from single-generation AI applications to complex task automation.
DeepMind CEO Hassabis: I sleep six hours a day, usually feel energetic around 1 a.m.
Fortune recently featured an interview with Demis Hassabis, CEO of Google DeepMind, revealing his unconventional approach to rest and productivity. Hassabis disclosed that he sleeps very little, structuring his waking hours into two distinct work blo
OpenAI, Anthropic Vie for Market Share Despite Revenue Shortfalls
Despite recent reports suggesting OpenAI missed revenue targets, creating pressure on tech stocks this Tuesday, private AI lab investors remain resilient. Seasoned backers have confirmed they will not reduce investment despite negative media coverage





Home






