logo
RunComfy
  • ComfyUI
  • TrainerNew
  • Models
  • API
  • Pricing
discord logo
MODELS
Explore
All Models
LIBRARY
Generations
MODEL APIS
API Docs
API Keys
ACCOUNT
Usage

Seedance 2.5 Reference to Video: Consistent Cinematic Clips from References on Models and API | RunComfy

bytedance/seedance-2.5/reference-to-video/720p

Seedance 2.5 Reference to Video 720p blends up to 9 images, 3 videos, and 3 audio references with a prompt into a 4–30-second 720p clip, with optional native audio.

Text prompt for the video (Chinese ~≤500 characters, English ~≤1000 words recommended).
Image 1
Reference images to steer environment and style. Up to 9 images.
Reference video clips for camera motion and rhythm. Up to 3 clips.
Reference audio for mood and pacing. Up to 3 files.
Aspect ratio of the generated video.
The duration of the generated video in seconds. Between 4 and 30.
When true, the model outputs video with synchronized audio (speech, SFX, music).
Idle
The rate is $0.306 per second without a reference video, and $0.187 per second with a reference video. Without a reference video: total price = output video duration × $0.306. With a reference video: total price = (input reference video duration + output video duration) × $0.187. Example: 10s reference video + 10s output = 20s × $0.187 = $3.74. Image and audio references are not billed.

Introduction To Seedance 2.5 Reference to Video

Seedance 2.5 Reference to Video 720p turns reference images, short video clips, and audio into a cinematic 4–30-second clip at 720p that keeps your subject, wardrobe, and style consistent. Combine up to 9 images, 3 videos, and 3 audio files with a text prompt: images steer identity and style, videos carry camera motion and rhythm, and audio sets the mood. Optional synchronized native audio can be generated with the video. For developers, Seedance 2.5 on RunComfy runs in the browser and via an HTTP API, so you don't need to host or scale the model yourself.

Why Choose Seedance 2.5 Reference to Video 720p#


Seedance 2.5 Reference to Video 720p uses your reference material to drive generation at 720p. Blend up to 9 images, 3 videos, and 3 audio files into a single guided generation: images steer identity and style, videos carry camera motion and rhythm, and audio sets the mood — all combined via one text prompt, with optional synchronized native audio.


AdvantageWhat it means for you
Multi-reference controlCombine up to 9 images, 3 videos, and 3 audio files so several sources guide one coherent result.
Stronger consistencyReference images plus a clear prompt help anchor identity, wardrobe, and tone across frames.
Native audio in one passGenerate synchronized speech, effects, and music with the clip, or turn audio off for silent video.
Up to 30-second clipsDirect a longer single shot with steadier quality than stitched short takes.

Best Use Cases#


  • Consistent character videos: Keep a person, mascot, or product on-model across a shot.
  • Product references to cinematic clips: Turn product and environment references into a directed scene.
  • Style-locked brand videos: Carry a look, palette, or motion feel across variants.
  • Previsualization: Combine references and prompt to test a scene before a full shoot.

How It Works#


  1. Add your references: Upload reference images under Images (up to 9); add short reference Videos (up to 3) or Audio (up to 3) to guide motion or sound.
  2. Describe the shot: Write what should happen and how the camera behaves (subject action, push-in, pan).
  3. Respect the limits: Reference videos and audio should be about 2–15 seconds each; keep reference audio under 15 MB.
  4. Set aspect ratio, duration, and audio: Choose an aspect ratio and a 4–30-second duration, and decide whether to generate audio. Output resolution is fixed at 720p.
  5. Generate and refine: Swap references, refine the prompt, then generate again.

Parameters#


The table below lists the controls exposed by the Seedance 2.5 Reference to Video 720p tool on this page.


ParameterRequiredTypeDefaultRange / OptionsHow to choose
prompt*Yes (*)StringExample promptChinese ~≤500 characters or English ~≤1000 words recommendedDescribe the action and camera; references anchor identity, motion, and mood.
imagesNoArray (image URLs)Example imageup to 9jpeg, png, webp, bmp, tiff, gif; steer identity and style.
videosNoArray (video URLs)[]up to 3mp4, mov; ~2–15 s each; carry camera motion and rhythm.
audiosNoArray (audio URLs)[]up to 3wav, mp3; ~2–15 s, under 15 MB; set the mood.
aspect_ratioNoString16:916:9, 9:16, 1:1, 4:3, 3:4, 21:9, adaptiveMatch the destination frame; adaptive lets the model pick the closest ratio.
durationNoInteger54–30 seconds, in 1-second stepsShort clips for a single action; longer only when the prompt has a clear arc.
generate_audioNoBooleantruetrue / falseLeave on for synchronized speech, effects, and music; turn off for silent video.

\* Required field. Only the prompt is required; references are optional but recommended for consistent results.


Pricing#


The rate is $0.306 per second without a reference video, and $0.187 per second with a reference video.


  • Without a reference video: total price = output video duration × $0.306.
  • With a reference video: total price = (input reference video duration + output video duration) × $0.187. Example: 10s reference video + 10s output = 20s × $0.187 = $3.74.
  • Image and audio references are not billed.

Prompting Tips & Examples#


  • Let references anchor, prompt direct: Use images for what must stay stable; use the prompt for action and camera.
  • Keep clips short: Reference videos and audio around 2–15 seconds each; audio under 15 MB.
  • Name sound sources: State who speaks, what makes each sound, and the ambience.
  • Use negative instructions: State what you do not want (for example, no text, no watermark).

Improved prompt example


> Animate the reference into a cinematic sci-fi shot: the astronaut walks forward across the alien dunes as wind lifts glowing dust, the two pale moons rising, volumetric god rays sweeping across the landscape, the camera slowly pulls back to reveal a vast otherworldly desert, low ambient wind and a deep cinematic drone.


More Seedance 2.5 Pages to Try#


  • Seedance 2.5 Reference-to-Video (480p): For cheaper, faster drafts, use the 480p reference-to-video page.
  • Seedance 2.5 Text-to-Video: Generate from a prompt only, with no references.
  • Seedance 2.5 Image-to-Video: Animate a single still image.
  • Seedance 2.5 First & Last Frame: Bridge a start and end frame into a smooth transition.

Related Models

flux-3/keyframes-to-video

Generate video from multi-keyframe stills with optional audio

hailuo-02/pro/text-to-video

Generate sharp HD videos from text with Minimax Hailuo 02 Pro.

wan-2-5/text-to-video

Generate videos from text prompts with audio using Wan 2.5 Preview.

video-background-removal/fast/video-to-video

AI-powered tool for fast video-to-video backdrop swaps with pro-level precision.

minimax-h3/text-to-video

MiniMax H3: 768p/2K text-to-video with native stereo audio

kling-2-1/standard/image-to-video

Animate a single image into a smooth video with Kling 2.1 Standard.

Frequently Asked Questions

What is Seedance 2.5 Reference to Video 720p best used for?

It guides a 720p clip with reference images and optional video or audio, keeping identity, wardrobe, and style consistent. It fits consistent-character videos, product-reference clips, and style-locked brand videos.

How does reference-to-video generation work here?

You attach reference material — up to 9 images, 3 short videos, and 3 audio clips — and describe the shot in the prompt. References anchor identity, wardrobe, style, motion, and sound, while the prompt guides action and camera. Only the prompt is strictly required.

What reference limits and duration does Seedance 2.5 Reference to Video 720p support?

Up to 9 reference images, 3 reference videos, and 3 reference audio files (reference videos and audio about 2–15 seconds each; audio under 15 MB). Duration is a whole number of seconds from 4 to 30 (default 5). Aspect ratio can be 16:9 (default), 9:16, 1:1, 4:3, 3:4, 21:9, or adaptive. Output is fixed at 720p.

Do I need video and audio references to use it?

No. It can run from a text prompt plus images alone. Add short reference videos or audio when you want stronger motion or mood guidance.

Does Seedance 2.5 Reference to Video 720p generate audio?

Yes. generate_audio is on by default, so the model can output synchronized speech, effects, and music. Turn it off when you only need silent video.

Can developers call Seedance 2.5 Reference to Video 720p through the RunComfy API?

Yes. Prototype in the RunComfy model UI, then call the same template through the API with matching fields (prompt, images, videos, audios, aspect_ratio, duration, generate_audio). Generations consume credits on both paths.

How much does Seedance 2.5 Reference to Video 720p cost on RunComfy?

Without a reference video, total price = output video duration × $0.306. With a reference video, total price = (input reference video duration + output video duration) × $0.187. Example: 10s reference video + 10s output = 20s × $0.187 = $3.74. Image and audio references are not billed.

How does it differ from the 480p page?

Both share the same reference-guided path and 4–30-second window; this page outputs at 720p. For cheaper, faster drafts, use the 480p reference-to-video page.

Follow us
  • LinkedIn
  • Facebook
  • Instagram
  • Twitter
Support
  • Discord
  • Email
  • System Status
  • Affiliate
Video Models
  • Runway Aleph 2
  • Wan 2.5
  • Wan 2.6 Flash
  • MiniMax H3 Max
  • Wan 3.0 Reference To Video
  • Seedance 2.5 Reference to Video 480p
  • View All Models →
Image Models
  • Wan 2.6 Image to Image
  • Seedream 5.0 Pro
  • Flux 2 Klein 9B
  • Nano Banana Pro
  • Nano Banana 2 Edit
  • Z Image Turbo LoRA
  • View All Models →
Legal
  • Terms of Service
  • Privacy Policy
  • Cookie Policy
RunComfy
Copyright 2026 RunComfy. All Rights Reserved.

RunComfy is the premier ComfyUI platform, offering ComfyUI online environment and services, along with ComfyUI workflows featuring stunning visuals. RunComfy also provides AI Models, enabling artists to harness the latest AI tools to create incredible art.

Examples Of Seedance 2.5 Reference to Video

Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...