ByteDance's Seedance 2.5 Generates 30-Second Videos From 50 References

ByteDance's Seedance 2.5 lands on Runway with 30-second clips, 50 multimodal references, and built-in audio — but there's a resolution catch worth knowing.

·
·
AuthorRunway
Read2 min
TopicVideo · Api
  • Seedance 2.5 is live on Runway — ByteDance's new video model now available on all paid Runway plans.
  • 30-second single-pass generation — doubles the 15-second cap of Seedance 2.0, with multi-round extension support.
  • 50 multimodal references per generation — 30 images, 10 video clips, 10 audio clips, up from 9 images and 3 clips in 2.0.
  • Resolution caveat — Seedance 2.5 on Runway outputs 480p/720p only; use Seedance 2.0 for 1080p output.
  • Cost — 30 credits/second at 720p; use Seedance 2.0 Fast/Mini for cheap iteration drafts.
  • Access — available in Custom mode, Workflows, and Agent on Runway.

Seedance 2.5, ByteDance's latest video generation model, is now available on Runway. The headline feature is straightforward: you can generate a single, coherent video clip up to 30 seconds long from a text prompt and up to 50 reference inputs spanning images, video, and audio. That's a meaningful jump from where AI video generation has been sitting.

There's a small asterisk on the "1080p" framing in Runway's announcement. Seedance 2.5 on Runway outputs at 480p and 720p. For 1080p, you use Seedance 2.0. The two models coexist on the platform and are optimized for different jobs , more on that below.

What actually changed from 2.0

Since Seedance 2.0, ByteDance noticed a shift in what users expect from video models: from generating a clip to completing a creative work. Seedance 2.5 builds on the unified multimodal audio-video architecture of 2.0, delivering major breakthroughs in long-form storytelling, multimodal reference, and editing. The three upgrades that matter most in practice:

  • Longer clips: Up to 30 seconds per generation in a single pass, with support for multiple rounds of extension. Seedance 2.0 caps at 15 seconds.
  • More references: Up to 50 multimodal reference inputs per generation , 30 images, 10 video clips, and 10 audio clips , far above Seedance 2.0's 9 images and 3 clips.

Keep reading

Don't miss what's next in AI

Join 300,000+ engineers and researchers who get the signal, not the noise. Create a free account to read the rest of this story.

  • Full access to in-depth AI research breakdowns
  • Be the first to know what's trending before it hits mainstream
  • Daily curated papers, repos, and industry moves