/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Sources: DeepSeek R2's launch delay is due to training issues on Huawei Ascend chips, prompting a switch to Nvidia chips for training and Huawei's for inference

Difficulties of training the start-up's latest system with Huawei's semiconductors highlight dependence on Nvidia

Financial Times

Context & Ripple Effects

DeepSeek's R2 delay follows earlier reporting that a shortage of Nvidia server chips in China had also constrained the project. The new account adds an execution constraint: Huawei hardware was being tested for training despite prior reports of gaps in its training performance, connectivity and software stack relative to Nvidia.

The reported split between Nvidia for training and Huawei for inference anticipates DeepSeek's later plan to use Ascend for smaller R2 variants while retaining Nvidia for its largest models.

First-order effects

  • R2's release is delayed as DeepSeek moves training workloads to Nvidia chips, reinforcing Nvidia as the immediate dependency for its largest-model development.
  • Huawei retains an inference role, but the reported training problems limit Ascend's use in the most demanding development workflow.

Second-order effects

  • DeepSeek must operate a mixed hardware stack, separating training from inference rather than standardizing on one supplier; this adds integration and deployment complexity.
  • The result sharpens pressure on Huawei to improve the training-side software, stability and inter-chip performance identified in earlier reporting on Ascend's training gaps.

Third-order effects

  • If this division persists, Chinese AI developers may treat domestic accelerators as a viable inference layer before relying on them for frontier-model training, creating a more durable split between smaller-model and largest-model hardware choices.
  • The episode shows that access to chips alone does not remove compute execution risk: software maturity and systems performance can remain the binding constraint even where alternate hardware is available.

The trend: AI developers are increasingly adopting heterogeneous compute strategies, assigning training and inference to different chip platforms as they balance performance, availability and software readiness.

Discussion

  • @realbobbyhealy @realbobbyhealy on x
    0 to 1 * Not https://www.ft.com/...
  • @alecolarizi @alecolarizi on x
    Huawei sent a team of engineers to DeepSeek's office to help the company use its AI chip to develop the R2 model, according to two people. Yet despite having the team on site, DeepSeek could not conduct a successful training run on the Ascend chip
  • @hsu_steve Steve Hsu on x
    DeepSeek R2 delay due to transition to Huawei Ascend chip for training? DS + HW engineers collaborating on CUDA to CANN migration is ultimately positive for HW in the long run. R2 release was originally expected last May. Since then at least one SOTA Chinese model has been
  • @dennisw5 Dennis Wilder on x
    So much for the vaunted Huawei and the myth that China can easily replace NVIDIA chips. DeepSeek's next AI model delayed by attempt to use Chinese chips https://www.ft.com/... via @FT
  • @ft @ft on x
    The Chinese AI start-up encountered persistent technical issues during its R2 training process using Huawei's Ascend chips, highlighting the limits of Beijing's push to replace US technology https://www.ft.com/... [image]
  • @chinabeigebook @chinabeigebook on x
    “DeepSeek has delayed the release of its new model after failing to train it using #Huawei's chips, highlighting the limits of Beijing's push to replace 🇺🇲 tech” And what does this suggest the right USG policy response should be? https://www.ft.com/...
  • r/China_irl r on reddit
    DeepSeek新模型发布因华为芯片问题而推迟
  • r/LocalLLaMA r on reddit
    DeepSeek's next AI model delayed by attempt to use Chinese chips
  • r/singularity r on reddit
    Deepseek delayed R2 model as they “could not conduct a successful training run” on Huawei chips