Huawei this week outlined its AI accelerator roadmap, revealing significant performance gains for the upcoming Ascend 960, 970, and 980 neural processing units (NPUs) over the existing Ascend 910C and 950-series. The Ascend 960DT and 960PR, targeted for 2027, are expected to deliver FP8 training performance of 2 PFLOPS and FP4 inference performance of 4 PFLOPS and 8 PFLOPS, respectively. Successors Ascend 970 and Ascend 980 are projected to reach FP4 performance of 14 PFLOPS and 28 PFLOPS in subsequent years.
Speaking on the sidelines of the Huawei Connect conference, rotating chairman Eric Xu told Reuters: “Since we do not have enough capacity to even satisfy the demand in China, we do not have a plan to expand into the international market in a fully-fledged way.” He added that Huawei supplies limited volumes to “some countries where demand is particularly strong,” without elaborating.
The company’s Atlas clusters, which can scale to 120 EFLOPS, use optical networking to link up to 15,488 chips, positioning Huawei to compete with Nvidia’s GPU-based systems. However, the capacity constraints mean that even as Huawei’s Ascend 960-series may rival offerings from Nvidia and AMD, the vast majority of the world will remain reliant on Western suppliers for the foreseeable future.