Black Forest Labs Announces FLUX 3 Video, First Details Cover Generation Capabilities
Black Forest Labs has published the first part of its FLUX 3 Video release notes, focused on the model's generation capabilities. Full technical specifications, pricing, and benchmark data have not yet been disclosed.
Black Forest Labs has begun rolling out documentation for FLUX 3 Video, a new video generation model, starting with a post titled "FLUX 3 Video, Part 1: Generation." The release marks the company's continued expansion beyond image generation, where its FLUX.1 family has become a widely used open and commercial alternative to models from Stability AI and Midjourney.
As of publication, Black Forest Labs has not released full technical specifications for FLUX 3 Video. No context window size, parameter count, pricing per 1M tokens, or independent benchmark scores have been disclosed. The "Part 1" framing suggests this is the opening entry in a multi-part rollout, with additional posts likely to cover editing, control, or other capabilities in subsequent releases — following a pattern the company has used for prior FLUX launches.
What is confirmed
The only verified fact at this stage is the existence of a model called FLUX 3 Video and that Black Forest Labs is treating "generation" as a distinct, dedicated capability worth its own documentation. This implies the model supports text-to-video or image-to-video synthesis, consistent with the broader video generation market that includes OpenAI's Sora, Google DeepMind's Veo, and Runway's Gen series.
Beyond this, specific claims about output resolution, video length, frame rate, or generation speed have not been provided in the source material. Any performance figures Black Forest Labs shares in follow-up posts should be treated as company claims until independently verified by third parties.
Context: FLUX's trajectory
Black Forest Labs, founded by former Stability AI researchers, built its reputation on the FLUX.1 image models, which shipped in multiple tiers (schnell, dev, pro) with varying licensing terms and inference costs. A jump to "FLUX 3" alongside a dedicated video variant indicates the company is consolidating its next-generation architecture across both image and video modalities rather than treating video as a separate product line.
The absence of pricing and benchmark data at this stage is typical for staged releases, where technical reports precede API availability or open-weight downloads. Readers should expect Black Forest Labs to publish inference cost, hardware requirements, and comparative benchmarks (likely against Sora, Veo, or Kling) in later parts of this series.
What this means
This is a documentation-stage announcement, not a full product launch. There is no way yet to evaluate FLUX 3 Video's actual quality, cost, or availability against competitors. The practical significance will depend entirely on what Black Forest Labs discloses in the promised follow-up parts — particularly whether the model is open-weight (as FLUX.1 dev and schnell were) or closed and API-only, and how its pricing compares to existing video generation services. Until then, this announcement functions primarily as a signal that Black Forest Labs is moving into video generation as a core product line rather than a side experiment.
Related Articles
MiniMax H3 Becomes First Open Video Model to Top an AI Video Ranking
MiniMax has released open weights for H3, a 33-billion-parameter video model that ranks first in Video Editing and second in Text-to-Video on Artificial Analysis — the first time an open model has topped a video generation category. The model accepts text, images, video, and audio in a single prompt, though its highest-resolution module remains closed.
MiniMax Releases H3, a 33B-Parameter Omni-Modal Model That Generates 2K Video With Native Stereo Audio
MiniMax has published MiniMax-H3, a 33-billion-parameter omni-modal generative model capable of producing up to 15 seconds of 2K video with native stereo audio. The model accepts text, image, video, and audio inputs, though its full 2K pipeline depends on a hosted preprocessing component not included in the open-source release.
ByteDance's Seedance 2.5 Generates 30-Second AI Video Clips With Synced Audio
ByteDance released Seedance 2.5, an AI video model that generates synchronized video and audio in a single pass, producing clips up to 30 seconds long that can be extended further. That's roughly triple the length of Google's Gemini Omni Flash.
Liquid AI Releases LFM2.5-2.6B, a 2.6B-Parameter Agentic Model with 128K Context for On-Device Use
Liquid AI has released LFM2.5-2.6B, a 2.6B-parameter model trained on 34 trillion tokens with a 128K context window, built for on-device agentic workloads. The company claims it is competitive with models four times its size on tool use and instruction following.
Comments
Loading...