DeepSeek V4

7 articles tagged with DeepSeek V4

August 21, 2026
model releaseDeepSeek

DeepSeek Releases V4 Flash Vision Exp, an Experimental Multimodal MoE Model with 1M Context

DeepSeek has released V4 Flash Vision Exp, an experimental vision-enabled variant of DeepSeek V4 Flash 0731 that adds image understanding while matching the base model's text performance. The sparse mixture-of-experts model uses 13B active parameters out of 284B total and supports a 1M token context window.

August 14, 2026
changelogDeepSeek

DeepSeek to Quadruple API Prices for V4 Pro and V4 Flash Starting August 16

DeepSeek will raise API output token pricing roughly fourfold starting August 16, introducing peak and off-peak rates for its V4 Pro and V4 Flash models. Despite the increase, DeepSeek remains cheaper than competitors like OpenAI's GPT-5.6 Sol and Moonshot's Kimi K3.

August 12, 2026
model releaseDeepSeek

DeepSeek Releases V4 Pro 0813 With 1.05M Token Context Window, Priced at $0.43/M Input

DeepSeek has shipped the general availability release of DeepSeek V4 Pro, codenamed 0813, featuring a 1,049,000-token context window. The mixture-of-experts model is priced at $0.43 per million input tokens and $0.87 per million output tokens, and is live now on OpenRouter.

model releaseNVIDIA

Nvidia Reportedly Building Trillion-Parameter Nemotron 4 to Match Chinese Open Models

Nvidia is reportedly building Nemotron 4, an open-weight model with at least one trillion parameters — double the size of Nemotron 3 Ultra. The company has tripled its cloud spending on in-house training to $28 billion through 2031, with an earliest possible release this fall.

August 1, 2026
changelogDeepSeek

DeepSeek Launches 'V4 Flash Latest' Alias with 1M+ Token Context on OpenRouter

DeepSeek has published a new routing endpoint, deepseek-v4-flash-latest, that always points to the newest model in its V4 Flash family. The endpoint offers a 1,049K token context window and pricing of $0.09/M input and $0.18/M output tokens via OpenRouter.

July 31, 2026
model releaseDeepSeek

DeepSeek Releases V4 Flash 0731: 1M-Token MoE Model at $0.14/M Input Tokens

DeepSeek has released V4 Flash 0731, a sparse mixture-of-experts model with 13B active parameters out of 284B total and a 1049K token context window. The model targets coding, reasoning, and agent workflows, priced at $0.14 per million input tokens and $0.28 per million output tokens via OpenRouter.

July 23, 2026
model release+1

Laguna S 2.1 Launches: Startup Claims Cheaper-Than-DeepSeek Pricing and Better-Than-V4-Pro Performance

A new Western AI lab has released Laguna S 2.1, a model that Reddit users and early testers describe as cheaper than DeepSeek V4 Flash while outperforming DeepSeek V4 Pro. Pricing, benchmark scores, and context window details remain undisclosed as of publication.