CONNECTED THINKING

Knowledge atlas

Follow people, viewpoints and their original evidence.

1 people · 1 sources · 1 viewpoints

IN CONTEXT

AI video input flexibility

Choose a viewpoint. Follow it back to the conversation.

Multi-modal input composition

IN CONTEXTAI video input flexibility1 viewpoints
2026-04-15

Equal sectors are reading positions, not rankings.

Showing 1–1 of 1 viewpoints · Newest sources first

1 / 1

Selected viewpoint

Multi-modal input composition

Seedance 2.0 accepts up to 9 images, 3 video clips, 3 audio files, and a text prompt simultaneously—assigning distinct creative roles to each: composition from images, camera movement from video, rhythm from audio, and descriptive intent from text.

These are individual perspectives, not a measure of consensus. Source material stays in its original language.

Supporting evidence

How to make remarkable videos with Seedance 2.0 – Replicate blog

Original excerpt

Most video models take a text prompt and give you a clip. Seedance 2.0 works differently. You can feed it up to 9 images, 3 video clips, 3 audio files, and a text prompt. The model understands how to use each piece. You can pull the composition from a photo, the camera movement from a video clip, the rhythm from an audio track, and describe how it all works together in words.

Publication dates describe the sources, not changes in belief.