model releaseByteDance

ByteDance's Seedance 2.5 Generates 30-Second AI Video Clips With Synced Audio

TL;DR

ByteDance released Seedance 2.5, an AI video model that generates synchronized video and audio in a single pass, producing clips up to 30 seconds long that can be extended further. That's roughly triple the length of Google's Gemini Omni Flash.

2 min read
0

ByteDance has released Seedance 2.5, an updated version of its AI video generation model that produces synchronized video and audio in a single generation pass. The model creates clips up to 30 seconds long, which can be extended multiple times for longer sequences.

By comparison, Google's Gemini Omni Flash tops out at roughly 10 seconds per clip. ByteDance has not disclosed pricing or API rate limits for Seedance 2.5.

What's new

Seedance 2.5 accepts a broader set of reference inputs than its predecessor. Users can upload up to 30 images, 10 video clips, and 10 audio files as references, allowing the model to construct scenes with multiple characters and varying camera angles from those inputs. ByteDance also says textures, lighting, and skin detail have improved over Seedance 2.0, though the company has not published quantitative benchmark comparisons for these visual quality claims.

To demonstrate the model, ByteDance produced a short film titled "The Missing Pair" using Seedance 2.5 exclusively, showcasing multi-character scenes and camera work generated by the system.

Availability

Seedance 2.5 is live now on ByteDance's Jimeng AI and Doubao Pro platforms. API access through BytePlus ModelArk is planned but not yet available, so third-party developers cannot currently integrate the model into external products.

Track record

The previous generation, Seedance 2.0, currently leads the image-to-video leaderboard among models with audio generation, according to independent benchmarking site Artificial Analysis. That model gained attention in the film community after "District 9" director Neill Blomkamp used it to create "Nightborne," a 13-minute short film generated entirely with AI — one of the longer AI-generated narrative works produced to date.

What this means

The 30-second ceiling matters less for raw duration than for what it enables: fewer stitching seams. Combining separately generated clips into a coherent scene is one of the persistent pain points in AI video production, since lighting, character consistency, and audio sync tend to drift between segments. A single-pass 30-second generation with native audio reduces how many of those seams a production team has to manually patch.

The expanded reference inputs — 30 images, 10 videos, 10 audio files — point toward ByteDance targeting professional ad and content studios rather than casual users experimenting with text prompts. That's a different competitive lane than OpenAI's Sora or Runway's consumer-facing tools, and it puts Seedance closer to being a production tool than a novelty generator.

Still, the comparison to Gemini Omni Flash's 10-second limit should be read carefully: longer native generation often trades off against per-second visual fidelity or coherence, and ByteDance has not released independent benchmark data to confirm quality holds steady across the full 30 seconds. Until BytePlus ModelArk API access ships and outside benchmarking firms test the model directly, ByteDance's specific quality claims — improved textures, lighting, skin detail — remain unverified. The Artificial Analysis leaderboard placement for Seedance 2.0 is independently confirmed, which lends some credibility to ByteDance's trajectory, but 2.5 has not yet been through the same third-party evaluation.

Related Articles

model release

OpenAI Reportedly Developing 'Astra' Model Family for Multi-Day Autonomous Problem-Solving

OpenAI is reportedly developing a new model family called Astra, designed to coordinate multiple agents on complex problems over hours or days. The models are already in testing and would be first to go through a planned U.S. government pre-release review, according to The Information.

model release

Google DeepMind Launches Gemini Robotics 2, a Single VLA Model for Arms to Humanoids

Google DeepMind has introduced Gemini Robotics 2, a vision-language-action model it calls its most advanced yet, designed to control everything from tabletop robot arms to full-body humanoids. The company also released Gemini Robotics ER 2, an embodied reasoning model that replaces ER 1.6.

model release

Thinking Machines Releases Inkling Small, a 12B-Active-Parameter Model That Beats Its Larger Predecessor on Key Benchmar

Thinking Machines has released Inkling Small, an open-weights reasoning model with 276 billion total parameters but only 12 billion active. According to Artificial Analysis, it scores nearly as high as the company's larger Inkling model while using roughly a third of the parameters and far fewer output tokens per task.

model release

DeepSeek Releases V4-Flash-0731, a 284B-Parameter Model That Beats Its Own Larger Pro Variant on Agentic Benchmarks

DeepSeek has shipped the full release of DeepSeek-V4-Flash-0731, a 284B-parameter model that according to DeepSeek outperforms its own larger V4-Pro (Preview) on agentic and coding benchmarks. Unsloth has published quantized GGUF versions, with lossless 8-bit weights requiring 162GB of storage.

Comments

Loading...