Infery.ai
← All models

Seedance 2.5 Reference to Video

seedance-2.5-reference-to-video

Video generationby bytedance

Dreamina Seedance 2.5 generates video from up to 50 multimodal references images, video, audio, and style inputs, locking a character, set, and palette across a full 30-second take for production-grade consistency.

Example

Details

Accepts
text + audio + image + video

Pricing

Input
2675 cr / 1M tokens

Prices in credits (1 credit = $0.01).

Data schema

Input

FieldTypeDescription
promptrequiredstringThe text prompt used to generate the video.
durationstringenum: auto, 4, 5, 6, 7, 8Duration of the video in seconds. Supports 4 to 30 seconds, or auto to let the model decide based on the prompt.
audio_urlsarrayReference audio to guide video generation. Refer to them in the prompt as @Audio1, @Audio2, etc. Supported formats: MP3, WAV. Up to 10 files. Each file must be 1.8 to 30.2 seconds and no larger than 15 MB; combined duration must not exceed 30.2 seconds. If audio is provided, at least one reference i…
image_urlsarrayReference images to guide video generation. Refer to them in the prompt as @Image1, @Image2, etc. Supported formats: JPG, PNG, WebP, BMP, TIFF, GIF, HEIC, HEIF. Max 30 MB per image. Up to 30 images. Total files across all modalities must not exceed 50.
resolutionstringenum: 480p, 720pVideo resolution - 480p for faster generation, 720p for balance.
video_urlsarrayReference videos to guide video generation. Refer to them in the prompt as @Video1, @Video2, etc. Supported formats: MP4, MOV. Up to 10 videos. Each video must be 1.8 to 30.2 seconds and no larger than 200 MB; combined duration must not exceed 30.2 seconds. Dimensions must be 300 to 6,000 pixels per…
end_user_idThe unique user ID of the end user.
aspect_ratiostringenum: auto, 21:9, 16:9, 4:3, 1:1, 3:4The aspect ratio of the generated video. Use 16:9 for landscape, 9:16 for portrait/vertical, 1:1 for square, 21:9 for ultrawide cinematic, or auto to let the model decide.
generate_audiobooleanWhether to generate synchronized audio for the video, including sound effects, ambient sounds, and lip-synced speech. The cost of video generation is the same regardless of whether audio is generated or not.

Output

FieldTypeDescription
seedintegerThe seed used for generation.
videoThe generated video file.