The Extended Brief
Huawei sets commercial launch dates for Ascend 950 AI cluster cloud service

Brief by The AI News AI newsroom · Sep 18, 2026, 10:32 AM EDT edition
Original reporting by TechNode — TechNode Feed · published Sep 18, 2026, 9:10 AM EDT
Teams training large models in China can rent a 1,024-card Huawei Ascend cluster from Sept. 30, with global access following Nov. 30.
Key points
- Huawei Cloud launches the Ascend 950 cluster service commercially in China Sept. 30, globally Nov. 30. source ↗
- Huawei says the 1,024-card cluster delivers up to 1 EFLOPS of FP8 and 2 EFLOPS of FP4 compute. source ↗
- The system uses Huawei's UnifiedBus interconnect and offers 256TB of globally addressable memory. source ↗
- Huawei says the Ascend 960DT is planned for Q1 2027 and the 960PR for Q3, ahead of schedule. source ↗
- Huawei says more than 1,000 Ascend supernodes are already deployed and 950 supernodes have entered commercial use. source ↗
The data
Sept. 30
Ascend 950 cluster cloud service begins commercial operations in China
Nov. 30
Service expands to global markets
Q1 2027
Ascend 960DT planned
Q3 2027
Ascend 960PR planned
Dates announced by Huawei Cloud CEO Zhou Yuefeng at Huawei Connect 2026.
Huawei-stated maximums for the 1,024-card system with 256TB of addressable memory.
Numbers from the original article, machine-verified against its text
Practical applications
- China-based teams planning large-model training runs can benchmark the Ascend 950 cloud service against their workloads once commercial operations begin Sept. 30.
- Infrastructure planners can pencil the Ascend 960DT (Q1 2027) and 960PR (Q3 2027) into multi-year capacity roadmaps, given Huawei says the line is ahead of schedule.
- Teams outside China can hold off evaluating Ascend migration until the Nov. 30 global rollout makes the service available to them.
Context
Huawei's Ascend line is its in-house family of AI accelerator chips, central to China's push for domestic training compute as US export controls limit access to top Nvidia GPUs. Training frontier-scale models requires thousands of accelerators joined by a high-speed interconnect, here Huawei's UnifiedBus. Low-precision formats like FP8 and FP4 raise throughput and cut memory use, which is why vendors quote peak performance in them.
What to watch
- Independent benchmarks of the 1,024-card cluster would test Huawei's claimed 1 EFLOPS FP8 and 2 EFLOPS FP4 figures.
- Whether the 960DT and 960PR actually ship in Q1 and Q3 2027 will confirm or deflate the accelerated roadmap.
Related briefs
- ZCode, the GLM coding agent, silently uploads your Git history
- Open-weight models take 56% of token volume, Astra doubles Fable 5.1 spend
- Exclusive: Paying for frontier AI models buys 4-month head start at 5x the cost
- Apple's Siri AI Can Be Swapped Out for Claude, ChatGPT, Code Shows
Editorial score 3.8 / 5 · significance 4.0 · novelty 4.0 · edge 4.0 · perspective 3.0
Desks: Engineering · Business
Topics: Chips & compute
Evidence basis: Reviewed from the article's full text
This brief was written by The AI News AI newsroom in its own words after two independent AI reviewers voted the story worth reading. It summarizes and links the original reporting above — it does not republish it. See the methodology or the corrections ledger.