ByteDance/seedance-v2.0-mini-reference-to-video
seedance-v2.0-mini-reference-to-video

Generate reference-guided videos with Seedance 2.0 Mini using images, videos, and audio, optional sound, flexible duration, and affordable ByteDance output.

Reference to Video

Related Seedance 2.0 Mini Models

Seedance 2.0 Mini Reference to Video
Add audio
Upload from device
mp3, wav
0/3 Audio
Upload Images
Click or drag to upload
Supported formats: jpg, png, webp
0/9 images
Add video
Upload from device
Supported formats: mp4, webm, mov
0/3 Video
prompt
Translate
s
Try the AI Video Generator now

README

Seedance V2.0 Mini Reference-to-Video API for Affordable Multimodal Video Generation

Seedance V2.0 Mini Reference-to-Video is ByteDance's cost-effective model for generating videos from image, video, and audio references. It combines multimodal guidance with flexible output controls, making it suitable for creators and developers who need consistent visual direction without the higher-resolution output of the standard Seedance V2.0 model.

Key Features of Seedance V2.0 Mini Reference-to-Video

  • Multimodal Reference Guidance: Guide a generation with images, videos, and audio clips in one request.
  • Flexible Reference Inputs: Use 1-9 images, up to 3 videos, and up to 3 audio clips. Images may be omitted when at least one reference video is supplied.
  • Reference-Aware Prompting: Address uploaded media directly in the prompt with numbered image, video, and audio placeholders.
  • Practical Output Resolutions: Generate at 480p or 720p for cost-conscious production workflows.
  • Flexible Video Duration: Create clips from 4 to 15 seconds.
  • Multiple Aspect Ratios: Choose from 21:9, 16:9, 9:16, 1:1, 4:3, and 3:4.
  • Optional Generated Sound: Sound is enabled by default and can be disabled when a silent result is preferred.
  • No Seed Parameter: Seedance V2.0 Mini models do not support seed control.

How to Use Seedance V2.0 Mini Reference-to-Video

  1. Write a Prompt: Describe the intended scene, motion, camera behavior, timing, and relationship between the supplied references.
  2. Add Visual References: Upload at least one image or video. You can provide 1-9 images and 0-3 videos; an empty image list is valid only when a reference video is present.
  3. Add Optional Audio References: Upload up to 3 audio clips when sound, rhythm, speech, or timing should influence the output.
  4. Reference Media in the Prompt: Use numbered placeholders such as image 1, video 1, and audio 1 through the API placeholder syntax shown in the API documentation.
  5. Configure the Output: Select 480p or 720p, a duration from 4-15 seconds, and one of the six supported aspect ratios.
  6. Choose Whether to Generate Sound: Keep sound enabled by default or disable it for silent video generation.
  7. Submit the Request: Send the request to the video task API and use the returned task identifier to check generation status.

When one or more reference videos are uploaded, the generation charge is doubled once. The multiplier is the same regardless of the number or duration of the uploaded reference videos.

Best Use Cases for Seedance V2.0 Mini Reference-to-Video

  • Character and Product Continuity: Use reference images or clips to preserve recognizable visual direction across new scenes.
  • Social Media Video: Produce vertical, square, landscape, or cinematic clips for common publishing formats.
  • Motion and Camera Guidance: Use a reference video to communicate movement, pacing, composition, or camera behavior.
  • Audio-Guided Creation: Align generated visuals with supplied music, speech, sound effects, or timing cues.
  • Rapid Creative Exploration: Test multimodal concepts at 480p or 720p before committing to higher-resolution production.

Note: Every request must include at least one image or video reference. Each uploaded video and audio clip must be between 2 and 15 seconds long. Audio-only requests are not supported.

Seedance V2.0 Mini Reference-to-Video vs Competitors

  • Compared with Seedance V2.0 Reference-to-Video: Both support image, video, and audio guidance, while the Mini model focuses on affordable 480p and 720p output instead of the standard model's higher-resolution options.
  • Compared with Wan 2.7 Reference-to-Video: Seedance V2.0 Mini supports separate image, video, and audio reference arrays and generated sound, while omitting seed control.
  • Compared with Vidu Q3 Reference-to-Video: Seedance V2.0 Mini extends reference guidance beyond images by accepting video and audio inputs.
  • Compared with Seedance V2.0 Mini Text-to-Video: Text-to-Video generates from a prompt alone, while Reference-to-Video adds image, video, and optional audio inputs for multimodal guidance.
  • Compared with Runway's Video Models: Seedance V2.0 Mini provides a direct reference-to-video API for teams that want image, video, and audio guidance in one request.

Explore multiple Creative Tools with Flaq AI

Explore several AI creation tools for quick image and video workflows in your browser, then scale successful ideas with Flaq AI's production-ready model APIs. Flaq AI provides a unified API layer for all models, making it easy to use and scale your workflows.