Create expressive video up to 30 seconds with picture and sound generated together.
Up to 30s per generationNative synchronized audioUp to 50 reference assets480p / 720p direct outputText, image, video and audio input
MODEL OVERVIEW
What is Seedance 2.5?
Seedance 2.5 is ByteDance's next-generation AI video model for creating longer,
more coherent scenes from prompts and multimodal references.
It combines multi-shot storytelling, stronger subject continuity, and native
audio generation in one workflow—giving creators more control from concept to output.
01
Longer ideas
Develop a complete sequence instead of assembling disconnected short clips.
02
Stronger direction
Guide character, style, framing, motion, and atmosphere with reference assets.
03
Picture and sound
Generate visuals, dialogue, ambience, and effects as one coordinated result.
01 · CINEMATIC SHOWCASES
Three distinct worlds, brought to life
Move from traditional stage performance and modern music to a lantern-lit city beneath
fireworks—each with its own atmosphere, movement, and visual direction.
Three cinematic directions — swipe to explore
CINEMATIC PERFORMANCE
Opera in motion
Traditional performance shaped through expressive movement, costume, and cinematic light.
PERFORMANCE STORY
From backstage to the spotlight
A performer moves from preparation to the stage in one connected visual journey.
ATMOSPHERIC WORLD
A city beneath fireworks
A lantern-lit historical world unfolds across a sequence of cinematic shots.
02 · PRECISE REFERENCE CONTROL
Turn references into creative direction
Combine up to 50 reference assets—including images, video, and audio—to communicate
the intended character, composition, visual style, movement, and atmosphere.
Use references to keep products recognizable, characters consistent, and each shot
closer to the original creative direction.
CharacterStyleFramingMotionSound
03 · NATIVE AUDIO AND PERFORMANCE
Create picture and sound as one experience
Generate visuals and audio together instead of treating sound as a separate step.
Dialogue, ambience, effects, and performance can develop alongside the image.
Create narrative content, social campaigns, and product stories where sound is part
of the idea from the beginning.
04 · SCENE TRANSFORMATION
Carry the same characters into a new world
Move a scene from a sunlit park to a vineyard while preserving the characters,
their relationship, and the natural rhythm of their movement.
Use reference-led transformation to explore new settings without losing the identity
and visual continuity that connect the sequence.
HOW TO USE SEEDANCE 2.5
From creative brief to generated video in three steps
01
Describe the scene
Write the subject, action, camera direction, visual style, and desired sound.
02
Add references
Provide images, video, or audio to guide characters, products, motion, and tone.
03
Generate and refine
Create the video, review the result, and adjust the prompt or references as needed.
BUILT FOR DEVELOPERS
Bring Seedance 2.5 into your product through one API
Simple asynchronous workflowUp to 50 multimodal references480p and 720p direct outputUsage-based pricing
Seedance 2.5 is ByteDance's AI video generation model for longer multi-shot scenes, multimodal reference control, stronger continuity, and native audio.
What inputs can I use?
You can create from text and guide the result with image, video, and audio references. The LSF integration supports up to 50 reference assets per request.
How long and what resolution can the output be?
Seedance 2.5 can generate video up to 30 seconds. Direct LSF output is currently available at 480p and 720p; use LSF Video Enhancement when you need a separate 1080p or 4K upscale workflow.
Does Seedance 2.5 generate audio?
Yes. It can generate picture and synchronized sound together, including dialogue, ambience, and effects guided by your prompt and references.
BRING YOUR IDEAS TO LIFE
Bring your next story to life with Seedance 2.5
Explore richer motion, native audio, and consistent multi-shot storytelling.