# Seedance 2.5 Reference to Video > ByteDance Seedance 2.5 reference-to-video via Fal.AI. Generates video from up to 50 multimodal references — up to 30 images, 10 videos, and 10 audio files — locking a character, set, and palette across a full 30-second take. Reference inputs are addressed from the prompt as @Image1, @Video1, @Audio1, and drive motion transfer, editing, extension, and lip-sync. - **Provider**: fal - **Model ID**: bytedance/seedance-2.5/reference-to-video - **Category**: video_generation - **Credits**: 5500 per request - **Pricing Type**: input_based ## API Endpoint Base URL: https://api.core.today/v1 ### Create Prediction POST /predictions ### Get Status GET /predictions/{job_id} ### Cancel DELETE /predictions/{job_id} ## Authentication Header: `X-API-Key: YOUR_API_KEY` ## Input Parameters - `duration` (string, **required**): Duration of the video in seconds (4-30). Required: upstream also accepts "auto", but the resulting length is unknown at request time and cannot be billed accurately, so this gateway requires an explicit duration. (Default: `5`; Options: `4`, `5`, `6`, `7`, `8`, `9`, `10`, `11`, `12`, `13`, `14`, `15`, `16`, `17`, `18`, `19`, `20`, `21`, `22`, `23`, `24`, `25`, `26`, `27`, `28`, `29`, `30`) - `aspect_ratio` (string, optional): The aspect ratio of the generated video. Use 16:9 for landscape, 9:16 for portrait/vertical, 1:1 for square, 21:9 for ultrawide cinematic, or auto to let the model decide. (Default: `auto`; Options: `auto`, `21:9`, `16:9`, `4:3`, `1:1`, `3:4`, `9:16`) - `audio_urls` (array, optional): Reference audio to guide video generation, addressed in the prompt as @Audio1, @Audio2, etc. Supported formats: MP3, WAV. Up to 10 files; each 1.8-30.2s and at most 15 MB, combined duration at most 30.2s. Requires at least one reference image or video. - `image_urls` (array, optional): Reference images to guide video generation, addressed in the prompt as @Image1, @Image2, etc. Supported formats: JPG, PNG, WebP, BMP, TIFF, GIF, HEIC, HEIF. Max 30 MB per image, up to 30 images. Total files across all modalities must not exceed 50. - `video_urls` (array, optional): Reference videos to guide video generation, addressed in the prompt as @Video1, @Video2, etc. Supported formats: MP4, MOV. Up to 10 videos; each 1.8-30.2s and at most 200 MB, combined duration at most 30.2s. Supplying any reference video switches billing to the video_in rate, which bills the reference footage as well as the output. - `generate_audio` (boolean, optional): Whether to generate synchronized audio for the video, including sound effects, ambient sounds, and lip-synced speech. The cost of video generation is the same regardless of whether audio is generated. (Default: `True`) - `prompt` (string, **required**): The text prompt used to generate the video. Address reference inputs as @Image1, @Video1, @Audio1, etc. For dialogue, put the spoken words in double quotes. - `resolution` (string, optional): Video resolution — 480p for faster generation, 720p for balance. (Default: `720p`; Options: `480p`, `720p`) ## Example Request ```json { "model": "bytedance/seedance-2.5/reference-to-video", "input": { "duration": "5", "image_urls": [ "https://v3b.fal.media/files/b/0a8eba37/Cqg-4Uwzyz4DELfceT1CF_a17e588773ec45b1a9e6f100a787b80b.jpg" ], "aspect_ratio": "auto", "generate_audio": true, "prompt": "An octopus finds a football in the ocean and excitedly calls its octopus friends to come and play. Cut scene to an octopus football game under the sea.", "resolution": "720p" } } ``` ## Response Format ```json { "job_id": "abc123", "status": "pending", "provider": "replicate", "model": "black-forest-labs/flux-schnell", "created_at": "2026-01-01T00:00:00Z", "result": null, "error": null } ``` Status values: `pending`, `processing`, `completed`, `failed`, `cancelled` ## Usage Flow 1. POST /predictions with model and input → receive job_id 2. Poll GET /predictions/{job_id} until status is `completed` or `failed` 3. Result contains output URL(s) or data ## Output Type json ## Tags bytedance, seedance, video-generation, reference-to-video, multimodal-reference, character-consistency, video-editing, native-audio, lip-sync ## Documentation https://fal.ai/models/bytedance/seedance-2.5/reference-to-video