Realtime AI News
Alibaba Cloud's Zhenwu M890 Supernode Adapts Qwen3.8, Goes Live on Bailian for Inference
Alibaba's Zhenwu M890 supernode has successfully adapted Qwen3.8, Alibaba's flagship model with 2.4 trillion parameters, and is now available on the Bailian platform for inference services. The supernode interconnects 64 Zhenwu M890 chips via ICN Switch 1.0 at 800GB/s, making it the first supernode in China to run a model exceeding 2 trillion parameters.
On July 23, Alibaba announced that its Zhenwu M890 supernode has successfully adapted Qwen3.8 and is now live on the Alibaba Cloud Bailian platform for model inference — the first supernode in China to run a large model with over 2 trillion parameters, as reported by QbitAI.
Qwen3.8 is Alibaba's latest flagship model with 2.4 trillion parameters. Running a model of this scale requires distributing parameters across dozens or even hundreds of high-speed interconnected GPU cards for high-throughput, low-latency collaborative computation. Conventional AI clusters, constrained by limited communication bandwidth, cannot meet these demands.
The Zhenwu M890 is Pingtouge's next-generation training-inference integrated AI chip, natively supporting precision levels from FP32 to FP4, covering everything from high-precision training to ultra-low-precision inference. The supernode connects 64 Zhenwu M890 chips through the ICN Switch 1.0 interconnect, achieving 800GB/s chip-to-chip bandwidth with 9TB of total memory — sufficient for expert parallelism in a 2-trillion-parameter MoE model.
On the software side, Alibaba implemented full-stack collaborative optimization from the chip layer to the cloud platform for Qwen models. This allows Qwen3.8 to run smoothly on the Zhenwu supernode while significantly improving inference efficiency and reducing costs. Test data shows up to 1.5x inference performance improvement in Agentic reasoning scenarios through operator optimization and hardware-software co-design.
The Zhenwu M890 supernode's adaptation of Qwen3.8 demonstrates that China's domestic AI chip ecosystem is maturing rapidly. Pingtouge's transition from early training-focused chips to commercial deployment in ultra-large-scale inference scenarios marks a critical step forward for domestic AI hardware in the premium market segment.
For Alibaba Cloud, connecting the Zhenwu supernode with the Bailian platform creates a full-stack loop from chip to model to cloud service. This means Alibaba Cloud not only provides compute power but has deeply integrated capabilities across the chip, interconnect, and platform layers, building a higher competitive moat.
With Qwen3.8's inference service now officially live on Bailian, enterprise users can directly call this domestic ultra-large-scale model, running on homegrown chip infrastructure. This provides a self-controllable option for the large-scale deployment of domestic AI applications.
Why it matters
The Zhenwu supernode successfully running a 2.4-trillion-parameter model proves domestic AI chips can handle ultra-large-scale inference, and Alibaba Cloud's full-stack loop from chip to model is reshaping the AI infrastructure competition landscape.
Nearby Updates
All07/23, 14:38
China releases first AI agent interconnection standard system, creating digital 'IDs' for AI agents
China has released its first AI agent interconnection standard system, establishing unified identity authentication and communication protocols for AI agents from different vendors. The standard aims to solve interoperability challenges that have been a bottleneck for enterprise AI agent deployment.
07/23, 14:09
ServiceNow invests $40M in Indian banking AI software firm BusinessNext
ServiceNow has invested $40 million in Indian banking software specialist BusinessNext at a $700 million valuation. The deal gives ServiceNow a strategic partner to expand its AI-powered financial services offerings globally.
07/23, 13:00
Tianpule Large Model V4.7 Released with Enhanced AI Music Remixing and Cover Capabilities
Quwan Technology released Tianpule 4.7 at WAIC 2026, focusing on making AI-generated music more controllable and editable. The upgrade introduces enhanced Remix and Cover features, allowing users to modify style, arrangement, and vocals while preserving original melodies.
07/23, 12:42
Qujing Technology establishes East China HQ in Hangzhou, plans 10,000-GPU AI Token factory
Qujing Technology has established its East China regional headquarters in Hangzhou's Qianjiang Century City. The company plans to build a 10,000-GPU-card AI Token factory within five years to serve surging demand for large model training and inference compute.