August 22, 2026

AIincider

AI News. No Noise. Just Signal.

Huawei’s Ascend 950DT Debuts With Homegrown HBM

3 min read
Huawei's Ascend 950DT AI training chip is live on Huawei Cloud with 144GB of in-house HBM, its boldest push yet to cut Nvidia out. Read the full breakdown.

Huawei has put its newest AI training accelerator, the Ascend 950DT, onto Huawei Cloud this month, moving the launch forward from a schedule that had pointed at the fourth quarter. The detail that matters most is not the raw compute number. It is the memory: the chip carries 144GB of high bandwidth memory that Huawei designed and sourced in house.

Why Memory Is the Real Bottleneck

High bandwidth memory, or HBM, is the stacked DRAM that sits next to an AI accelerator and feeds it data. Without enough of it, and enough bandwidth to move data in and out, a fast processor spends most of its time waiting. HBM supply is the tightest link in the entire AI hardware chain, and roughly 80 to 88 percent of global capacity sits with two South Korean firms, SK hynix and Samsung.

Export controls have kept the newest Nvidia parts, and the HBM that goes with them, out of reach for Chinese buyers. That has left Chinese labs training on older silicon or on whatever they could stockpile. Huawei’s answer has been to build the whole stack domestically, memory included.

What the 950DT Brings

Huawei Cloud vice president Chen Lin confirmed the August cloud debut at the company’s INSPIRE Creators event, according to Huawei Central. The 950DT is aimed at training and decoding, the memory hungry end of the workload spectrum. It pairs 2 PFLOPS of FP8 compute with 144GB of Huawei’s own HiZQ 2.0 HBM, 4TB per second of memory bandwidth, and a 2TB per second interconnect that is roughly 2.5 times what the older Ascend 910C offered.

Full commercial availability is still slated for the fourth quarter, when the chip is also expected to anchor the Atlas 950 SuperPoD. That system is designed to link up to 8,192 of these accelerators for around 8 EFLOPS of FP8 performance. Chinese model builders are the obvious first customers, with DeepSeek widely tipped as an early adopter.

Why It Matters

Cloud access before general sale is a sensible way to get scarce silicon in front of the customers who can actually use it, and it lets Huawei prove the software stack while volume ramps. The harder question is whether in house HBM can be produced at the yields and volumes that training runs demand, which is a manufacturing problem more than a design one.

If it works, China gets a training path that does not route through Korean memory or American accelerators. Watch the Q4 SuperPoD deployments and any independent benchmarks from the labs actually running on this hardware.

Continue Reading…

Leave a Reply