DeepSeek releases V4 preview, claims parity with GPT-4o and Claude 3.5 Sonnet
DeepSeek released a preview of its V4 model on April 24, 2026, claiming the open-source system matches leading closed-source models from Anthropic, Google, and OpenAI. The company emphasized improved coding capabilities and compatibility with domestic Huawei chips, but did not disclose training costs or hardware specifications.
DeepSeek V4 Preview Released with Claimed Parity to Leading US Models
Chinese AI company DeepSeek released a preview of its V4 model on April 24, 2026, claiming the open-source system can compete with leading closed-source models from Anthropic, Google, and OpenAI.
Key Details
DeepSeek did not disclose:
- Training costs for V4
- Hardware used for training
- Specific benchmark scores
- Context window size
- Pricing information
The company emphasized that V4 represents a "major improvement" over prior models, particularly in coding capabilities. DeepSeek explicitly highlighted compatibility with domestic Huawei technology, marking what the company describes as a milestone for China's chip industry.
Context and Controversy
The V4 preview arrives approximately one year after DeepSeek's R1 model disrupted the US AI industry. DeepSeek claimed R1 was trained at a fraction of the cost of leading American systems, though specific cost figures were never independently verified.
US officials have accused DeepSeek of using banned Nvidia chips for training. Separately, Anthropic has claimed DeepSeek misused Claude to improve its own products, though details of these allegations remain unclear.
Coding Focus
According to DeepSeek, V4's enhanced coding performance targets capabilities that have become central to AI agents and driven adoption of systems like ChatGPT Codex and Claude Code. The company did not provide specific benchmarks comparing V4's coding performance to competitors.
What This Means
DeepSeek's V4 preview continues the pattern established with R1: bold claims about competitive performance without disclosed training costs, hardware specifications, or independent benchmark verification. The emphasis on Huawei chip compatibility suggests China's AI industry is working to reduce dependence on restricted Western semiconductor technology, though the practical performance implications remain unclear. Until DeepSeek releases concrete benchmarks and technical details, the actual capabilities of V4 relative to GPT-4o, Gemini, and Claude 3.5 Sonnet cannot be independently assessed.
Related Articles
Moonshot AI's Kimi k3 claims top performance among Chinese models with 1M token context
Moonshot AI has released Kimi k3, positioning it as China's leading AI model. The company claims the model features a 1 million token context window and improved reasoning capabilities, though independent benchmarks are not yet available.
Moonshot AI releases 2.8T parameter Kimi K3, pricing at $3/$15 per million tokens
Chinese AI lab Moonshot AI released Kimi K3, a 2.8 trillion parameter model priced at $3 per million input tokens and $15 per million output tokens. The model is currently available via API, with open weights promised by July 27, 2026. This represents the most expensive pricing from a Chinese AI lab to date, matching Anthropic's Claude Sonnet series.
Google delays Gemini 3.5 Pro release after disappointing coding performance in June training update
Google has delayed the release of Gemini 3.5 Pro past its June deadline due to coding performance issues. The company retrained the model in late June with new data but saw disappointing results, according to Bloomberg. An upgraded Flash model is now in testing with partners.
Moonshot AI Releases Kimi K3: Open-Weight Multimodal Reasoning Model with 1M Context Window
Moonshot AI has released Kimi K3, an open-weight multimodal reasoning model with a 1-million token context window. The model is priced at $3 per 1M input tokens and $15 per 1M output tokens, available through OpenRouter.
Comments
Loading...