2026-10-11 16:37 UTC

A LocalLLaMA post claims DeepSeek now trains its models on Huawei Ascend 950 accelerators rather than NVIDIA hardware, and confirmation from DeepSeek or Huawei — or credible technical corroboration — would mark China's leading open-model lab's concrete shift off NVIDIA training silicon, while refutation marks another unverified echo.

state: corroboratedheat: highuncertainty: mediumconvergesscott: mediumdeepseek huawei-ascend ai-infrastructureDeepSeekHuaweiLiang Wenfeng
Surfaced 2026-10-02T06:49:29Z — Official DeepSeek announcement: "今天,我们正式开源面向华为昇腾算力平台的基础设施组件,涵盖 TileLang 高级语言编译工具、计算库、分布式通信库。所有组件与此前面向英伟达平台的开源组件一一对应。" ("Today we officially — Consolidating the Sep 30 origin-walk and grounding into state: the post traces (conf 0.95) to DeepSeek's own WeChat announcement of open-sourced Ascend training infrastructure with NVIDIA-parity components — first-party plus independent Huawei and Bloomberg-derived lines corroborate substantial Ascend migration (operators, post-training on 910Cs, serving) while contradicting the strong off-NVIDIA pre-training reading. Attention has decayed past peak (157/33, ~0 pts/h, 25th percentile); the magnitude-valve flag overstates spread since both evidence objects are the same WeChat content (gallery repost + echo reconstruction), so heat cools to low despite a hot deepseek neighborhood.

What is this?

DeepSeek is a Hangzhou-based Chinese lab widely described in the supplied coverage as China's leading open-weight model developer; its V4 family launched April 2026 and was reportedly optimized to run on Huawei's Ascend 950 processors rather than Nvidia. The specific 'now trains on Ascend 950' claim is layered and the sources partially conflict: a Huawei-led team completed full-parameter post-training (not pre-training) of the 1.6T-parameter V4-Pro on ~1,000 older Ascend 910C chips, Huawei said V4-Flash training used Ascend, and DeepSeek has ordered 160,000 Ascend 950DTs for an Inner Mongolia data center — but Bloomberg-derived reporting says that cluster is for inference and that DeepSeek still relies on Nvidia for pre-training, while a leaked, unconfirmed Liang Wenfeng transcript calls DeepSeek's ~16,000 Ascend 950 allocation insufficient to train its next frontier model. One outlet (AI in Asia) flatly claims V4-Pro was fully trained on Ascend 950 Supernode clusters, contradicting the more careful post-training-only accounts, and the training-capable 950DT reportedly only ships in Q4 2026 — so the post's claim is corroborated for post-training and serving, contradicted for full pre-training, and turns entirely on what 'trains' means.

Why it matters to Scott

Converges with the export-controls compute-chokepoint argument his wikis already carry, and the grounding's resolution sharpens it into a dated-receipts claim: Ascend substitution lands at post-training and serving (V4-Pro on ~1,000 910Cs, 160K 950DTs ordered for Inner Mongolia) while Bloomberg-derived reporting keeps full pre-training on NVIDIA — exactly the inference-vs-pre-training compute split he tracks in inference economics. As a local-inference practitioner who serves open weights on non-CUDA stacks, an Ascend-co-designed DeepSeek would also make backend portability of future weights a live question; relevance stays medium until DeepSeek or Huawei confirms and the 'trains' ambiguity resolves.
work:technology.large-language-modelsdev:technology.cudaip:concept.vendor-lock-indev:concept.hardware-aware-local-inferenceradar:concept.export-controlsradar:huawei-domestic-accelerator-priorityradar:deepseek-agi-over-commercializationradar:concept.inference-economicsradar:concept.model-trainingradar:concept.china
queries asked of Scott's wikis
  • open-weights strategy and model sovereignty
  • export controls compute chokepoint thesis China
  • local inference economics non-NVIDIA hardware
  • CUDA ecosystem lock-in alternative accelerator software stacks
  • inference vs pre-training compute split serving bottleneck
  • model-hardware co-design frontier open models

Measured heat

now 0 pts/hpeak 28 pts/hcomments 0/hpeers p0momentum: steady2 platformsage 290h
points/hour across evidence · reading as of 2026-10-12 02:59:37.977291+11:00 · deterministic, not a model opinion

How the heat travelled

09-29 14:00⭐ origin echo-reconstructedOfficial DeepSeek announcement: "今天,我们正式开源面向华为昇腾算力平台的基础设施组件,涵盖 TileLang 高级语言编译工具、计算库、分布式通信库。所有组件与此前面向英伟达平台的开源组件一一对应。" ("Today we officially
DeepSeek (杭州深度求索人工智能基础技术研究有限公司, via its official "DeepSeek" WeChat public account) on blog (echo) · attributed from reddit.post.1wtz1i3
—
09-30 07:58first on r/LocalLLaMA · published · +18.0hDeepSeek now trained on Ascend 950
WebAssemblyMan
—
09-30 07:58amplified on r/LocalLLaMA 👑reddit.post.1wtz1i3
WebAssemblyMan
peak 162 · 35 comments · 100% of case engagement
09-30 08:20our radar first saw it · +18.3hdiscovery anchor: reddit.post.1wtz1i3—
10-02 06:47reached heat=high · +64.8h · via queue+ledger——
pace: p73 vs 1188 stories at the 168h mark (now 290h old) — ahead of openai-collective-cyber-defense (1.0x), behind trump-ai-force-czar-pledge (1.0x)

Evidence (2) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 redditDeepSeek now trained on Ascend 950
LocalLLaMA
Retrieved article excerpt

Open article · Retrieved 2026-09-30T08:25:15.335339+00:00

# Prove your humanity

We’re committed to safety and security. But not for bots. Complete the challenge below and let us know you’re
a real person.

[Reddit, Inc. © "2026". All rights reserved.](https://www.redditinc.com/)

[User Agreement](https://www.reddit.com/help/useragreement)
[Privacy Policy](https://www.reddit.com/help/privacypolicy)
[Content Policy](https://www.reddit.com/help/contentpolicy)
[Help](https://support.reddithelp.com/hc/en-us)
WebAssemblyMan16235
🟧 echo.blog ⭐Official DeepSeek announcement: "今天,我们正式开源面向华为昇腾算力平台的基础设施组件,涵盖 TileLang 高级语言编译工具、计算库、分布式通信库。所有组件与此前面向英伟达平台的开源组件一一对应。" ("Today we officially DeepSeek (杭州深度求索人工智能基础技术研究有限公司, via its official "DeepSeek" WeChat public account)——

Interpretation history

Decision trace