Multimodal referencing
Combine text, images, clips, and audio references in one coherent creative direction.
Turn text and mixed references into coherent cinematic scenes while preserving character identity, motion logic, and physical realism.
Combine text, images, clips, and audio references in one coherent creative direction.
Generate movement with convincing weight, timing, interactions, and environmental response.
Keep characters, products, clothing, and locations recognizable across changing shots.
Write the action, subject, visual style, and camera direction you need.
Launch the generator with Seedance 2.0, 16:9, five seconds, and 480p preselected.
Create the video, review the result, and adjust your prompt for another version.
Seedance 2.0 can interpret mixed visual and motion references, giving creators more control than a text-only workflow.
Its spatial and temporal reasoning helps preserve faces, clothing, products, and environmental details between shots.
Commercial use depends on your current plan, the selected model terms, and your rights to all prompts, uploads, logos, audio, and reference materials. Third-party platform rules may also apply.

Start from a clear prompt and continue editing in the ReelFlow generation workspace.