DeepSeek Delays AI Model Due to Huawei Chip Issues
DeepSeek has delayed the release of its new AI model after encountering difficulties in training with Huawei's chips, according to the Financial Times. The Chinese startup initially attempted to use Huawei's Ascend processors for training its R2 model but faced persistent technical issues.
Encouraged by authorities to adopt Huawei's technology over Nvidia's, DeepSeek found the Ascend chips unsuitable for training, although they are still being used for inference. This setback has delayed the model's launch from May, causing the company to lose ground to competitors.
The challenges faced by DeepSeek underscore the limitations of Chinese chips compared to their U.S. counterparts, as the company continues to work with Huawei to resolve these issues. Despite Huawei's support, including sending engineers to assist, DeepSeek has yet to achieve a successful training run on the Ascend chips.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like Silicon Brief or Daily AI Brief.
Also, consider following us on social media:
More from Data Centers
Sep 30 Sophia Space and Redwire Explore Orbital Data Centers Sep 30 SKF Recreates Greta Garbo With AI for Magnetic Bearing Campaign Sep 30 LG Innotek Targets $5.94 Billion in Semiconductor and Physical AI Businesses Sep 29 Efficient Computer Raises $97 Million to Scale Its Processors Sep 29 Hikvision Adds HIKO AI Engine to Hik-Connect 7Silicon Brief
Weekly coverage of AI hardware developments including chips, GPUs, cloud platforms, and data center technology.
Whitepaper
Tensordyne Napier: What If One Rack Could Do the Work of Nine?
Tensordyne
This Tensordyne whitepaper presents Napier, an inference-focused AI processor and rack-scale system based on the company’s TDN Math logarithmic number system. It examines infrastructure requirements for large mixture-of-experts and agentic models, compares major inference architecture approaches, and details the TDN AIP processor, TDN72 pod, TDN Link fabric, and Napier Ultra configuration. The paper reports simulation-based performance, cost, and accuracy-validation results, including Tensordyne’s projected comparison of one Napier rack with a nine-rack Nvidia Rubin plus Groq deployment; the chip is reported as taped out and in fabrication.
Read moreYou may also like
US Agencies Accuse Chinese AI Firms of Industrial Scale Model Distillation
Chinese AI Agents Deceived Evaluators in Controlled Tests
Anthropic Attributes Its Largest Measured Distillation Campaign to Alibaba
European AI Firms Reject Calls to Slow Model Development
Huawei Upgrades Stellar AI Network for AI Computing
Daily AI Brief: the AI news that matters, in your inbox.