ByteDance multimodal video model · ByteDance Seed
Seedance 2.0 guide for multimodal AI video creation
A source reviewed guide to Seedance 2.0, including its input modes, audio, control, practical limits, and current Genflow settings.

The short answer
Seedance 2.0 is ByteDance Seed's unified multimodal audio video model. It accepts text, image, audio, and video references, and its official release describes up to 15 second multi shot output with synchronized dual channel audio, extension, and targeted editing.
When to choose it
Seedance 2.0 is a practical fit for short commercial clips that need text to video, image direction, or reference based generation. Move to Seedance 2.5 when the brief needs up to 30 seconds or a much larger reference set.
Official model facts
What the official sources say
Vendor documented capabilities, kept separate from Genflow product settings.
Four input modalities
The official release describes combined text, image, audio, and video input in one generation workflow.
Up to 15 second multi shot output
Seedance 2.0 is presented as producing 15 second audio video sequences with multiple shots and synchronized sound.
Extension and targeted editing
Prompts can guide continuation and changes to particular clips, characters, actions, or story elements.
Genflow
Available in Genflow
Read directly from the current Genflow model configuration.
- Generation
- Keyframe to video, Reference to video, Text to video
- Duration
- 4 to 15 seconds
- Aspect ratios
- 16:9, 4:3, 1:1, 3:4, 9:16, 21:9
- Output
- 480p, 720p, 1080p, 4k
- Native audio
- Yes
Where it fits
- Flexible entry points for text, images, video references, and sound references.
- Useful balance of short form duration, motion, camera direction, and native audio.
- Genflow exposes text to video, keyframe to video, and reference to video paths for the main 2.0 model.
What to review
- The official release notes remaining issues with multiple subject consistency, text rendering, complex editing, and occasional audio distortion.
- Fast and Mini are Genflow variants with a smaller set of generation modes and resolution choices than the main 2.0 entry.
Seedance 2.0
Practical use cases
Short product films
Create concise product movement, atmosphere, voice, and camera direction in one clip.
Reference based ad concepts
Combine visual, motion, storyboard, and sound references around a 15 second idea.
Fast creative testing
Use the Fast or Mini variants in Genflow when iteration speed matters more than the broadest controls.
FAQ
Questions creators ask
What is Seedance 2.0?
Seedance 2.0 is a multimodal audio video generation model from ByteDance Seed that accepts text, image, audio, and video references.
What is the difference between Seedance 2.0 and 2.5?
Seedance 2.5 extends single pass output from 15 to 30 seconds, supports a much larger reference set, and adds more precise controls. Seedance 2.0 remains useful for shorter work and offers more entry modes in Genflow today.
Does Genflow include Seedance 2.0 Fast and Mini?
Yes. Genflow lists the main Seedance 2.0 model plus Fast and Mini variants. Their available generation modes and output settings are narrower than the main model.
Primary sources
Facts on this page are checked against these official model sources.