On June 5 Huawei Cloud released ModelArts Next, positioning the platform from a "training platform" to an "Agent-native" training-inference base. The new version integrates Agent Runtime, Memory Service, and TokenHub, forming a complete "model + Agent" co-design stack. The official numbers show that the Agent Runtime frees 70% of idle compute, the Memory service cuts long-task token usage by 60%, and the TokenHub compute utilization rate is up 40%.