How many references can I use in a Dreamina video?
The limit depends on the selected model and generation mode. Dreamina's current public model pages describe up to 50 multimodal inputs for Seedance 2.5 and up to 12 reference resources for Seedance 2.0. These are capacity limits, not a recommendation to fill every slot.
The Seedance 2.0 page also lists image, video and audio categories. Do not add category counts together and assume the combined result overrides the overall 12-reference limit. An overall capacity does not mean that every file type can occupy every slot.
Which references do I actually need?
Add a file when it explains something important that the text prompt cannot describe precisely enough.
References improve the information available to the model. They do not guarantee identical faces, products, sound or motion.
How do I tell the model which reference does what?
Use the reference labels shown in your workspace and assign each one a role. For example, adapt this brief to your uploaded assets:
Use Image 1 for the product's appearance and Image 2 for the studio background. Use Video 1 only for the slow camera movement. Keep the product centered and recognizable. Do not introduce the person or props from Video 1.
This is an illustrative brief, not a tested output or a guarantee that a typed label creates a reference. Attach or select the assets in the product interface, then use the labels assigned there.
Can I combine image, video and audio references?
Yes, supported multimodal Seedance workflows combine different input types. Select the model and mode first, then upload the accepted assets. Model access, file requirements and credit cost are separate from the maximum number of references.
For a free trial, check that the chosen reference workflow is available within your account's current credits. A model's published capacity is not a promise that its largest setup can be completed on a free balance.
What should I do when references conflict or an upload fails?
Do not keep adding files without resolving the conflict. A smaller coherent reference set can communicate the task more clearly than several unrelated assets.
Can I use first and last frames, or start with text only?
When the beginning or ending composition matters, use a dedicated first-frame or first-and-last-frame mode if your selected model offers it. Choose that mode explicitly rather than assuming two pictures uploaded as general references will automatically become endpoints.
You can also start without references in a supported text-to-video workflow. Add a reference when a specific identity, movement, scene or sound needs tighter direction. Begin with the Seedance 2.5 model page, select the relevant workflow, and prepare only the assets that help explain the shot.