model releaseDeepSeek

Deepseek v4 launching on Huawei chips exclusively, signaling China's AI independence progress

TL;DR

Deepseek v4 is launching in the coming weeks running exclusively on Huawei chips, marking a major milestone in China's effort to reduce dependency on foreign semiconductors. Chinese tech giants including Alibaba, Bytedance, and Tencent have ordered hundreds of thousands of Huawei Ascend 950PR units to deploy the model through their cloud services.

2 min read
1

Deepseek v4 Launching Entirely on Huawei Chips

Deepesk v4 is expected to launch within weeks running entirely on Huawei's Ascend 950PR chips, according to reporting from The Information. The move represents a significant shift in China's AI infrastructure strategy, with the model receiving no early access review from Nvidia—only Chinese chip manufacturers got preview access.

Chip Performance and Demand Surge

Huawei claims the Ascend 950PR delivers approximately 2.8x the computing power of Nvidia's H20 chip, though it remains below the H200's performance. The chip reportedly commands a 20 percent price premium following massive orders from major Chinese tech companies.

Alibaba, Bytedance, and Tencent have collectively ordered hundreds of thousands of Ascend 950PR units to run Deepseek v4 through cloud services and integrate it into their own applications, according to five people familiar with the matter. This concentration of orders from China's largest tech firms signals confidence in both the model and domestic chip viability.

Development Partnership

Deepesk spent months collaborating with Huawei and chip designer Cambricon to port v4 to Chinese-made hardware. The effort reflects a broader strategy to decouple AI development from Western semiconductor supply chains, particularly following US export controls that have constrained chip availability for Chinese companies.

Huawei continues facing production bottlenecks stemming from these same export restrictions, though the surge in Deepseek v4 orders suggests immediate demand exceeds supply constraints.

What This Means

Deepesk v4's exclusive reliance on Huawei hardware marks a tangible outcome of China's multi-year push toward semiconductor self-sufficiency. The decision to exclude Nvidia from early access—a departure from industry norm—signals confidence in domestic alternatives and reduces dependency on external validation. The aggressive procurement by Alibaba, Bytedance, and Tencent indicates the AI market sees viable alternatives to Nvidia, though performance gaps remain. Sustained production constraints and the 20 percent price premium suggest China's chip ecosystem still faces scaling challenges despite technical progress.

Related Articles

model release

DeepSeek Releases Experimental V4-Flash-Vision-Exp, Claims Near-Parity With Opus 4.8 on Agent Benchmarks

DeepSeek has released V4-Flash-Vision-Exp, an experimental multimodal extension of V4-Flash that adds image understanding while preserving text reasoning capabilities. The company claims the model approaches or beats Anthropic's Opus 4.8 on its internal multimodal agent benchmarks.

model release

DeepSeek Releases V4 Flash Vision Exp, an Experimental Multimodal MoE Model with 1M Context

DeepSeek has released V4 Flash Vision Exp, an experimental vision-enabled variant of DeepSeek V4 Flash 0731 that adds image understanding while matching the base model's text performance. The sparse mixture-of-experts model uses 13B active parameters out of 284B total and supports a 1M token context window.

changelog

DeepSeek to Quadruple API Prices for V4 Pro and V4 Flash Starting August 16

DeepSeek will raise API output token pricing roughly fourfold starting August 16, introducing peak and off-peak rates for its V4 Pro and V4 Flash models. Despite the increase, DeepSeek remains cheaper than competitors like OpenAI's GPT-5.6 Sol and Moonshot's Kimi K3.

model release

Anonymous 'Ox Alpha' Reasoning Model Appears on OpenRouter with Free 1M-Token Context

A stealth model called Ox Alpha has appeared on OpenRouter, offering a 1 million token context window at no cost during its preview period. The model's developer remains anonymous, and OpenRouter says it is acting only as a router, not the model's owner or provider.

Comments

Loading...