DeepSeek Releases V4-Flash-Base: 292B Parameter Base Model
DeepSeek has released V4-Flash-Base, a 292 billion parameter base model now available on Hugging Face. The model uses BF16, I64, F32, and F8_E4M3 tensor types and is distributed in Safetensors format.
DeepSeek V4-Flash-Base: 292B Parameter Base Model Released
DeepSeek has released V4-Flash-Base, a 292 billion parameter base model now available on Hugging Face. The model represents the base version of DeepSeek's V4-Flash series.
Technical Specifications
The model contains 292 billion parameters and supports multiple tensor types: BF16 (bfloat16), I64 (64-bit integer), F32 (32-bit float), and F8_E4M3 (8-bit float in E4M3 format). Files are distributed in the Safetensors format, which provides safer serialization than traditional pickle-based formats.
Availability and Deployment
The model weights are available for download on Hugging Face as part of a collection containing 4 items. According to the Hugging Face listing, no inference providers currently support deployment of this model. The collection was last updated approximately 4 hours ago and has 307 downloads.
Missing Information
DeepSeek has not yet published a model card with detailed information about training data, benchmark performance, capabilities, or pricing. Context window size, training cutoff date, and specific use cases remain undisclosed. As a base model, V4-Flash-Base typically requires fine-tuning for specific tasks, unlike instruction-tuned variants.
What This Means
The release of a 292B parameter base model signals DeepSeek's continued development of large-scale models, though the lack of documentation makes technical evaluation impossible at this stage. The "Flash" designation suggests optimization for speed, consistent with other models in the industry using similar naming conventions. The use of F8_E4M3 tensor types indicates potential support for efficient inference through quantization. Without benchmark scores or a detailed model card, organizations should wait for complete documentation before considering deployment.
Related Articles
NVIDIA Releases Nemotron-3-Embed-1B-BF16: 1.14B Parameter Multilingual Embedding Model with 2048-Dimensional Vectors
NVIDIA has released Nemotron-3-Embed-1B-BF16, a 1.14 billion parameter text embedding model supporting 34 languages with a 32,768 token context window. The model generates 2048-dimensional embeddings and was derived from Ministral-3-3B-Instruct-2512 through two rounds of structured pruning and distillation, first to 2B then to 1.14B parameters.
Moonshot AI's Kimi k3 claims top performance among Chinese models with 1M token context
Moonshot AI has released Kimi k3, positioning it as China's leading AI model. The company claims the model features a 1 million token context window and improved reasoning capabilities, though independent benchmarks are not yet available.
Moonshot AI releases 2.8T parameter Kimi K3, pricing at $3/$15 per million tokens
Chinese AI lab Moonshot AI released Kimi K3, a 2.8 trillion parameter model priced at $3 per million input tokens and $15 per million output tokens. The model is currently available via API, with open weights promised by July 27, 2026. This represents the most expensive pricing from a Chinese AI lab to date, matching Anthropic's Claude Sonnet series.
NVIDIA Releases Cosmos 3 Edge: 4B-Parameter World Model for Real-Time Robot Control at 15 Hz
NVIDIA has released Cosmos 3 Edge, a 4-billion-parameter open world model designed for edge AI systems. The model delivers real-time robot control at 15 Hz on NVIDIA Jetson devices, generating 32 actions per inference at 640×360 resolution.
Comments
Loading...