DeepSeek releases V4 preview, claims parity with GPT-4o and Claude 3.5 Sonnet
DeepSeek released a preview of its V4 model on April 24, 2026, claiming the open-source system matches leading closed-source models from Anthropic, Google, and OpenAI. The company emphasized improved coding capabilities and compatibility with domestic Huawei chips, but did not disclose training costs or hardware specifications.
DeepSeek V4 Preview Released with Claimed Parity to Leading US Models
Chinese AI company DeepSeek released a preview of its V4 model on April 24, 2026, claiming the open-source system can compete with leading closed-source models from Anthropic, Google, and OpenAI.
Key Details
DeepSeek did not disclose:
- Training costs for V4
- Hardware used for training
- Specific benchmark scores
- Context window size
- Pricing information
The company emphasized that V4 represents a "major improvement" over prior models, particularly in coding capabilities. DeepSeek explicitly highlighted compatibility with domestic Huawei technology, marking what the company describes as a milestone for China's chip industry.
Context and Controversy
The V4 preview arrives approximately one year after DeepSeek's R1 model disrupted the US AI industry. DeepSeek claimed R1 was trained at a fraction of the cost of leading American systems, though specific cost figures were never independently verified.
US officials have accused DeepSeek of using banned Nvidia chips for training. Separately, Anthropic has claimed DeepSeek misused Claude to improve its own products, though details of these allegations remain unclear.
Coding Focus
According to DeepSeek, V4's enhanced coding performance targets capabilities that have become central to AI agents and driven adoption of systems like ChatGPT Codex and Claude Code. The company did not provide specific benchmarks comparing V4's coding performance to competitors.
What This Means
DeepSeek's V4 preview continues the pattern established with R1: bold claims about competitive performance without disclosed training costs, hardware specifications, or independent benchmark verification. The emphasis on Huawei chip compatibility suggests China's AI industry is working to reduce dependence on restricted Western semiconductor technology, though the practical performance implications remain unclear. Until DeepSeek releases concrete benchmarks and technical details, the actual capabilities of V4 relative to GPT-4o, Gemini, and Claude 3.5 Sonnet cannot be independently assessed.
Related Articles
DeepSeek V4.1-Flash Cuts KV Cache Memory by Up to 8x, Matches Opus 5 on Coding Benchmark
DeepSeek released V4.1-Flash, a 552-billion-parameter model built to slash the memory overhead of long-context AI agents. The model cuts GPU cache needs to roughly a quarter of its predecessor's and matches closed models from OpenAI and Anthropic on select coding benchmarks.
DeepSeek Launches V4.1 Flash: Low-Cost MoE Model Claims to Beat V4 Pro
DeepSeek has released V4.1 Flash, a sparse mixture-of-experts model priced at $0.30 per 1M input tokens and $1.20 per 1M output tokens with a 1 million token context window. DeepSeek claims the model exceeds the larger V4 Pro on performance, speed, and task completion time.
DeepSeek Releases V4.1-Flash: 552B MoE Model Cuts KV Cache to 890 Bytes Per Token
DeepSeek has released V4.1-Flash, a 552B-parameter multimodal Mixture-of-Experts model supporting 1M-token context and activating only 8B parameters during prefill. The model uses a new Causal Encoder-Decoder architecture and Compressed Sparse Attention 2 to cut global KV cache to 890 bytes per token, roughly a quarter of its predecessor.
OpenAI Launches GPT-Live-1 API for Full-Duplex Voice Apps That Talk and Listen Simultaneously
OpenAI has released GPT-Live-1 as a developer API, a speech model capable of full-duplex conversation—listening and talking simultaneously. It already powers ChatGPT's voice mode and costs $0.05 per minute, with benchmark scores showing sharp improvements over GPT-Realtime-2.1.
Comments
Loading...