Huawei has introduced the Peerium Computing Architecture, a new computing architecture designed for the AI era that enables processors at the million scale to work as one computer. The announcement, made in Shanghai, addresses the ever-growing demand for AI compute, according to the company.
The Peerium Computing Architecture achieves strong scaling to the million-processor level through nested parallelism, unified memory addressing, and peer interconnect. It breaks through the Turing paradigm with the introduction of Nested BSP (Nested Bulk Synchronous Parallel), extends the von Neumann single-machine architecture, and overturns the master–slave architecture that has prevailed for decades, so that a million processors truly become one larger computer.
UnifiedBus (UB) is the key interconnect technology that makes the Peerium Computing Architecture possible. Built on a single open protocol, UB is a high-speed bus that scales without limit to connect CPUs, NPUs, memory, SSDs, network interface cards (NICs), and switches. With UB, peer interconnect is achieved across compute, storage, and networking.
The Atlas 950 SuperPoD and SuperPoD-based SuperClusters are the first-generation product built on the Peerium Computing Architecture. An Atlas 950 SuperCluster with 256,000 cards is already being deployed, and the Atlas 960 system based on near-packaged optics (NPO) is currently under testing.
Eric Xu, Huawei’s Rotating Chairman, said, “In the AI era, Huawei is drawing on the Peerium Computing Architecture we pioneered to continuously build the SuperPoDs and SuperPod-based SuperClusters that meet customer needs for training and inference, making computing power available everywhere and intelligence accessible to all.”
The new architecture could have significant implications for the AI industry, which requires massive computational resources for training and inference. By enabling a million processors to function as a single computer, Huawei aims to provide scalable computing power that can support advanced AI models and applications. The deployment of the Atlas 950 SuperCluster and the testing of the Atlas 960 system indicate that the technology is moving toward practical implementation. As reported by Media OutReach Newswire, this development marks a step toward making computing power more accessible and intelligence available to all.
