On July 23, Alibaba Cloud's Zhenwu M890 supernode completed full adaptation for Qwen3.8, and opened inference services on the Bailian platform. This is the first supernode in China to successfully run a large model above 2T parameters — a milestone for the joint deployment of domestic supernodes and ultra-large-parameter models. The Zhenwu M890 is Alibaba's in-house 128-card supernode, first announced in May and positioned for Agent concurrent inference, focusing on high density, low latency, and multi-task parallelism. Qwen3.8 is the latest flagship preview of Alibaba's Qwen, with 2.4T parameters, positioned as the open-source counter to top closed-source models. The Zhenwu + Qwen3.8 combination means the hardware layer and the model layer no longer work in silos — from day one, the chip, network, memory scheduling, and inference framework have been jointly tuned for Qwen3.8's 2T+ parameters. Why does this matter? In the past, running ultra-large-parameter models in China meant either relying on NVIDIA H100/H200 clusters, or stacking many machines and cards, with high latency and low utilization. M890 packs 128 cards into one supernode domain, with interconnect bandwidth and memory-sharing efficiency far above traditional clusters; after adapting Qwen3.8, it means users can one-click on Bailian get an inference-as-a-service that can carry 2T+ parameters on a domestic stack, no longer constrained by overseas compute supply. From an industry angle, this also sends a clear signal: the top cloud vendors are no longer content with the model layer and the hardware layer each doing their own thing, but are redefining the supernode from a system perspective. In the Agent era, inference is no longer offline benchmarking, but a high-concurrency, low-latency, long-context multi-task battlefield — only when chip + network + model are jointly optimized can real TOPS and cost-effectiveness be delivered. It's foreseeable that Qwen3.8's official release, and the next flagships from Claude / Gemini / DeepSeek, will all quickly replicate this path of domestic supernode + open-source large model. When domestic compute + domestic models + domestic inference frameworks form a closed loop, the so-called China-version AI stack finally has its foundation.