Discussion about this post

User's avatar
John Saunders's avatar

Over here we can't imagine China doing anything other than trying to compete with us on our terms. What's actually been happening is China going its own way, integrating AI with their systems as opposed to financializing it to make big bux.

Dmitry Mintz's avatar

One line in the release is carrying a lot: performance "approaching the limits of the underlying hardware." That is measured against Ascend, not against the path it would replace. A kernel at the ceiling of the chip it runs on, and a workload that costs no more than it did on CUDA, are different claims.

The joint target named is a 128-card supernode on Ascend 950. Huawei described the Atlas 960E on 17 September at 4,096 NPUs under unified memory addressing in a single pod, and up to 512,000 across a SuperCluster. Nothing published so far says what a production port costs at either size — engineering time, performance delta, whether anything had to be retrained.

No posts

Ready for more?