Black Forest Labs Announces FLUX 3 Video, First Details Cover Generation Capabilities
Black Forest Labs has published the first part of its FLUX 3 Video release notes, focused on the model's generation capabilities. Full technical specifications, pricing, and benchmark data have not yet been disclosed.
Black Forest Labs has begun rolling out documentation for FLUX 3 Video, a new video generation model, starting with a post titled "FLUX 3 Video, Part 1: Generation." The release marks the company's continued expansion beyond image generation, where its FLUX.1 family has become a widely used open and commercial alternative to models from Stability AI and Midjourney.
As of publication, Black Forest Labs has not released full technical specifications for FLUX 3 Video. No context window size, parameter count, pricing per 1M tokens, or independent benchmark scores have been disclosed. The "Part 1" framing suggests this is the opening entry in a multi-part rollout, with additional posts likely to cover editing, control, or other capabilities in subsequent releases — following a pattern the company has used for prior FLUX launches.
What is confirmed
The only verified fact at this stage is the existence of a model called FLUX 3 Video and that Black Forest Labs is treating "generation" as a distinct, dedicated capability worth its own documentation. This implies the model supports text-to-video or image-to-video synthesis, consistent with the broader video generation market that includes OpenAI's Sora, Google DeepMind's Veo, and Runway's Gen series.
Beyond this, specific claims about output resolution, video length, frame rate, or generation speed have not been provided in the source material. Any performance figures Black Forest Labs shares in follow-up posts should be treated as company claims until independently verified by third parties.
Context: FLUX's trajectory
Black Forest Labs, founded by former Stability AI researchers, built its reputation on the FLUX.1 image models, which shipped in multiple tiers (schnell, dev, pro) with varying licensing terms and inference costs. A jump to "FLUX 3" alongside a dedicated video variant indicates the company is consolidating its next-generation architecture across both image and video modalities rather than treating video as a separate product line.
The absence of pricing and benchmark data at this stage is typical for staged releases, where technical reports precede API availability or open-weight downloads. Readers should expect Black Forest Labs to publish inference cost, hardware requirements, and comparative benchmarks (likely against Sora, Veo, or Kling) in later parts of this series.
What this means
This is a documentation-stage announcement, not a full product launch. There is no way yet to evaluate FLUX 3 Video's actual quality, cost, or availability against competitors. The practical significance will depend entirely on what Black Forest Labs discloses in the promised follow-up parts — particularly whether the model is open-weight (as FLUX.1 dev and schnell were) or closed and API-only, and how its pricing compares to existing video generation services. Until then, this announcement functions primarily as a signal that Black Forest Labs is moving into video generation as a core product line rather than a side experiment.
Related Articles
PrismML Releases Ternary Bonsai 2 27B, a Compressed Reasoning Model with 262K Context
PrismML has released Ternary Bonsai 2 27B, a 27B-parameter reasoning model derived from Qwen3.8-27B that uses ternary weight compression to shrink to roughly 8.5 GB. The model supports a 262K-token context window, image understanding, tool calling, and thinks by default at 'xhigh' reasoning effort.
OpenAI RLHF Co-Inventor Launches Jev, a Non-LLM Model That Outputs Probabilities Instead of Text
TypeSafe AI, founded by RLHF co-inventor Diogo Almeida, has released Jev, a transformer-based model that outputs probabilities rather than text. Developers report it running 5 to 20 times cheaper and faster than LLMs for classification tasks.
Z.ai Releases GLM-5.3-FlashX, a 200 Tokens/Second Variant of Its GLM-5.3-Flash Model
Z.ai has released GLM-5.3-FlashX, a high-speed variant of GLM-5.3-Flash built on a hybrid sparse and linear attention architecture with 320B total parameters (18B active). The model supports a 1M-token context window and claims inference speeds of up to 200 tokens per second.
Moonshot AI's 2.8 Trillion-Parameter Kimi K3 Launches on Amazon Bedrock with 1M-Token Context
Moonshot AI's Kimi K3, described by the company as the first open model to reach 2.8 trillion parameters, is now available on Amazon Bedrock. It features native vision, a 1-million-token context window, and is the first open-weight model on Bedrock to support explicit prompt caching.
Comments
Loading...