Multimodal direction
Work from text, first or last frames and supported image references.
Model guide
A flexible route for text, frames and image references when the shot needs visual direction and native sound.
What it gives you
Seedance 2.5 combines longer duration control with first and last frames, multiple references and native audio. Soho exposes only the configurations it has mapped and priced.
Work from text, first or last frames and supported image references.
Choose from the active duration range instead of a single fixed length.
Generate with the model's supported audio behavior when enabled.
Model details
Text, first/last frame, and reference-guided clips from 4–30 seconds with selectable resolution and native audio.
Generated examples
Real Soho Frame examples identified by the model used.
Seedance 2.5
Seedance 2.5
Seedance 2.5
A clear workflow
Choose text, frames or references from the inputs the model actually accepts.
Select duration, resolution, ratio and audio where those controls are available.
The result returns beside its source with the generation settings attached.
Questions
The active Soho catalog currently documents up to 30 reference images. Frame modes and reference modes have different controls.
Yes. Native audio is supported in the mapped Soho route.
No. Source-video, source-audio and automatic-duration modes remain gated until their validation and pricing contracts are complete.
Open Video with the model selected and keep the result in its project.
Create with this model