Deepseek v4 launching on Huawei chips exclusively, signaling China's AI independence progress
Deepseek v4 is launching in the coming weeks running exclusively on Huawei chips, marking a major milestone in China's effort to reduce dependency on foreign semiconductors. Chinese tech giants including Alibaba, Bytedance, and Tencent have ordered hundreds of thousands of Huawei Ascend 950PR units to deploy the model through their cloud services.
Deepseek v4 Launching Entirely on Huawei Chips
Deepesk v4 is expected to launch within weeks running entirely on Huawei's Ascend 950PR chips, according to reporting from The Information. The move represents a significant shift in China's AI infrastructure strategy, with the model receiving no early access review from Nvidia—only Chinese chip manufacturers got preview access.
Chip Performance and Demand Surge
Huawei claims the Ascend 950PR delivers approximately 2.8x the computing power of Nvidia's H20 chip, though it remains below the H200's performance. The chip reportedly commands a 20 percent price premium following massive orders from major Chinese tech companies.
Alibaba, Bytedance, and Tencent have collectively ordered hundreds of thousands of Ascend 950PR units to run Deepseek v4 through cloud services and integrate it into their own applications, according to five people familiar with the matter. This concentration of orders from China's largest tech firms signals confidence in both the model and domestic chip viability.
Development Partnership
Deepesk spent months collaborating with Huawei and chip designer Cambricon to port v4 to Chinese-made hardware. The effort reflects a broader strategy to decouple AI development from Western semiconductor supply chains, particularly following US export controls that have constrained chip availability for Chinese companies.
Huawei continues facing production bottlenecks stemming from these same export restrictions, though the surge in Deepseek v4 orders suggests immediate demand exceeds supply constraints.
What This Means
Deepesk v4's exclusive reliance on Huawei hardware marks a tangible outcome of China's multi-year push toward semiconductor self-sufficiency. The decision to exclude Nvidia from early access—a departure from industry norm—signals confidence in domestic alternatives and reduces dependency on external validation. The aggressive procurement by Alibaba, Bytedance, and Tencent indicates the AI market sees viable alternatives to Nvidia, though performance gaps remain. Sustained production constraints and the 20 percent price premium suggest China's chip ecosystem still faces scaling challenges despite technical progress.
Related Articles
Reflection unveils 501B-parameter Beam, Mistral previews 1T-parameter Large 4, both open-weight
Reflection introduced Beam, a 501B-parameter mixture-of-experts model with 23B active parameters. Mistral said it is finishing Mistral Large 4, a 1T-parameter multimodal model with 49B active parameters. Both companies plan open-weight releases in October, and both are positioning the models against Chinese open-weight leaders.
OpenAI launches GPT-6 in ChatGPT with 'Intelligent UI' and interactive answers; Sol for paid users, Luna for free
OpenAI is rolling out GPT-6 to all ChatGPT tiers, with paying users on GPT-6 Sol and free users on GPT-6 Luna. The release adds 'Intelligent UI,' which renders answers as interactive charts, buttons, forms and mini apps, and lets the model respond while still thinking. OpenAI claims this cuts wait times by 44 percent.
Claude Haiku 5.5 arrives on Amazon Bedrock; Anthropic claims ~75% lower cost than Haiku 4.5
Claude Haiku 5.5 is available on Amazon Bedrock and Claude Platform on AWS. According to Anthropic, it is the fastest and most efficient model in the Claude 5.5 family and costs around 75% less than Claude Haiku 4.5 for most tasks. It is the first Haiku model with effort controls.
Anthropic releases Claude Haiku 5.5, claims ~75% lower running cost than Haiku 4.5
Anthropic released Claude Haiku 5.5 on October 7, 2026. The company claims it costs around 75% less to run than Haiku 4.5 and is its fastest model to date. Anthropic also halved Claude Sonnet 5.5's cache read pricing and added a monthly API credit for Max and Team subscribers.
Comments
Loading...