logo
RunComfy
  • ComfyUI
  • TrainerNew
  • Models
  • API
  • Pricing
discord logo
MODELS
Explore
All Models
LIBRARY
Generations
MODEL APIS
API Docs
API Keys
ACCOUNT
Usage

Seedance 2.5 Reference to Video: Consistent Cinematic Clips from References on Models and API | RunComfy

bytedance/seedance-2.5/reference-to-video/720p

Seedance 2.5 Reference to Video 720p blends up to 9 images, 3 videos, and 3 audio references with a prompt into a 4–30-second 720p clip, with optional native audio.

Text prompt for the video (Chinese ~≤500 characters, English ~≤1000 words recommended).
Image 1
Reference images to steer environment and style. Up to 9 images.
Reference video clips for camera motion and rhythm. Up to 3 clips.
Reference audio for mood and pacing. Up to 3 files.
Aspect ratio of the generated video.
The duration of the generated video in seconds. Between 4 and 30.
When true, the model outputs video with synchronized audio (speech, SFX, music).
Idle
Billed on total video seconds (input + output). Without reference videos the rate is $0.42/s of generated video; with reference videos the rate is $0.26/s for every counted second (reference video duration plus output duration). Image and audio references are not billed.

Introduction To Seedance 2.5 Reference to Video

Seedance 2.5 Reference to Video 720p turns reference images, short video clips, and audio into a cinematic 4–30-second clip at 720p that keeps your subject, wardrobe, and style consistent. Combine up to 9 images, 3 videos, and 3 audio files with a text prompt: images steer identity and style, videos carry camera motion and rhythm, and audio sets the mood. Optional synchronized native audio can be generated with the video. For developers, Seedance 2.5 on RunComfy runs in the browser and via an HTTP API, so you don't need to host or scale the model yourself.

Why Choose Seedance 2.5 Reference to Video 720p#


Seedance 2.5 Reference to Video 720p uses your reference material to drive generation at 720p. Blend up to 9 images, 3 videos, and 3 audio files into a single guided generation: images steer identity and style, videos carry camera motion and rhythm, and audio sets the mood — all combined via one text prompt, with optional synchronized native audio.


AdvantageWhat it means for you
Multi-reference controlCombine up to 9 images, 3 videos, and 3 audio files so several sources guide one coherent result.
Stronger consistencyReference images plus a clear prompt help anchor identity, wardrobe, and tone across frames.
Native audio in one passGenerate synchronized speech, effects, and music with the clip, or turn audio off for silent video.
Up to 30-second clipsDirect a longer single shot with steadier quality than stitched short takes.

Best Use Cases#


  • Consistent character videos: Keep a person, mascot, or product on-model across a shot.
  • Product references to cinematic clips: Turn product and environment references into a directed scene.
  • Style-locked brand videos: Carry a look, palette, or motion feel across variants.
  • Previsualization: Combine references and prompt to test a scene before a full shoot.

How It Works#


  1. Add your references: Upload reference images under Images (up to 9); add short reference Videos (up to 3) or Audio (up to 3) to guide motion or sound.
  2. Describe the shot: Write what should happen and how the camera behaves (subject action, push-in, pan).
  3. Respect the limits: Reference videos and audio should be about 2–15 seconds each; keep reference audio under 15 MB.
  4. Set aspect ratio, duration, and audio: Choose an aspect ratio and a 4–30-second duration, and decide whether to generate audio. Output resolution is fixed at 720p.
  5. Generate and refine: Swap references, refine the prompt, then generate again.

Parameters#


The table below lists the controls exposed by the Seedance 2.5 Reference to Video 720p tool on this page.


ParameterRequiredTypeDefaultRange / OptionsHow to choose
prompt*Yes (*)StringExample promptChinese ~≤500 characters or English ~≤1000 words recommendedDescribe the action and camera; references anchor identity, motion, and mood.
imagesNoArray (image URLs)Example imageup to 9jpeg, png, webp, bmp, tiff, gif; steer identity and style.
videosNoArray (video URLs)[]up to 3mp4, mov; ~2–15 s each; carry camera motion and rhythm.
audiosNoArray (audio URLs)[]up to 3wav, mp3; ~2–15 s, under 15 MB; set the mood.
aspect_ratioNoString16:916:9, 9:16, 1:1, 4:3, 3:4, 21:9, adaptiveMatch the destination frame; adaptive lets the model pick the closest ratio.
durationNoInteger54–30 seconds, in 1-second stepsShort clips for a single action; longer only when the prompt has a clear arc.
generate_audioNoBooleantruetrue / falseLeave on for synchronized speech, effects, and music; turn off for silent video.

\* Required field. Only the prompt is required; references are optional but recommended for consistent results.


Pricing#


Seedance 2.5 Reference to Video 720p bills on counted video seconds:


  • Without reference videos: $0.42 per second of generated video.
  • With reference videos: $0.26 per counted second (reference video duration plus output duration).
  • Image and audio references are not billed.

Prompting Tips & Examples#


  • Let references anchor, prompt direct: Use images for what must stay stable; use the prompt for action and camera.
  • Keep clips short: Reference videos and audio around 2–15 seconds each; audio under 15 MB.
  • Name sound sources: State who speaks, what makes each sound, and the ambience.
  • Use negative instructions: State what you do not want (for example, no text, no watermark).

Improved prompt example


> Animate the reference into a cinematic sci-fi shot: the astronaut walks forward across the alien dunes as wind lifts glowing dust, the two pale moons rising, volumetric god rays sweeping across the landscape, the camera slowly pulls back to reveal a vast otherworldly desert, low ambient wind and a deep cinematic drone.


More Seedance 2.5 Pages to Try#


  • Seedance 2.5 Reference-to-Video (480p): For cheaper, faster drafts, use the 480p reference-to-video page.
  • Seedance 2.5 Text-to-Video: Generate from a prompt only, with no references.
  • Seedance 2.5 Image-to-Video: Animate a single still image.
  • Seedance 2.5 First & Last Frame: Bridge a start and end frame into a smooth transition.

Related Models

minimax-h3-open/text-to-video

Open-weights text-to-video with 480p/768p and native stereo audio.

wan-2-2/speech-to-video

Turn photos into expressive videos with synced voice motion.

kling-video-o3/pro/image-to-video

Pro-tier image animation: 3-15s cinematic clips from $0.123 per second.

wan-2-2/fun-inpaint

Interpolates start-end frames with refined motion control presets

kling-video-o3/pro/reference-to-video

Pro-tier reference-driven 3-15s video generation from $0.19 per second.

veo-3/text-to-video

Generate premium-quality videos from text prompts with Google Veo 3.

Frequently Asked Questions

What is Seedance 2.5 Reference to Video 720p best used for?

It guides a 720p clip with reference images and optional video or audio, keeping identity, wardrobe, and style consistent. It fits consistent-character videos, product-reference clips, and style-locked brand videos.

How does reference-to-video generation work here?

You attach reference material — up to 9 images, 3 short videos, and 3 audio clips — and describe the shot in the prompt. References anchor identity, wardrobe, style, motion, and sound, while the prompt guides action and camera. Only the prompt is strictly required.

What reference limits and duration does Seedance 2.5 Reference to Video 720p support?

Up to 9 reference images, 3 reference videos, and 3 reference audio files (reference videos and audio about 2–15 seconds each; audio under 15 MB). Duration is a whole number of seconds from 4 to 30 (default 5). Aspect ratio can be 16:9 (default), 9:16, 1:1, 4:3, 3:4, 21:9, or adaptive. Output is fixed at 720p.

Do I need video and audio references to use it?

No. It can run from a text prompt plus images alone. Add short reference videos or audio when you want stronger motion or mood guidance.

Does Seedance 2.5 Reference to Video 720p generate audio?

Yes. generate_audio is on by default, so the model can output synchronized speech, effects, and music. Turn it off when you only need silent video.

Can developers call Seedance 2.5 Reference to Video 720p through the RunComfy API?

Yes. Prototype in the RunComfy model UI, then call the same template through the API with matching fields (prompt, images, videos, audios, aspect_ratio, duration, generate_audio). Generations consume credits on both paths.

How much does Seedance 2.5 Reference to Video 720p cost on RunComfy?

Without reference videos, billing is $0.42 per second of generated video. With reference videos, billing is $0.26 per counted second (reference video duration plus output duration); image and audio references are not billed.

How does it differ from the 480p page?

Both share the same reference-guided path and 4–30-second window; this page outputs at 720p. For cheaper, faster drafts, use the 480p reference-to-video page.

Follow us
  • LinkedIn
  • Facebook
  • Instagram
  • Twitter
Support
  • Discord
  • Email
  • System Status
  • Affiliate
Video Models
  • Seedance 2.5 Reference to Video 1080p
  • Seedance 2.5 1080p Text to video
  • Seedance 2.5 1080p
  • MiniMax H3 Open
  • Wan 2.6 Flash
  • Happy Horse 1.1 reference to video
  • View All Models →
Image Models
  • Qwen Image 3.0 Edit
  • Qwen Image 3.0 Pro Edit
  • Qwen Image 3.0
  • seedream 4.0
  • Flux 2 Flash Edit
  • Nano Banana Pro
  • View All Models →
Legal
  • Terms of Service
  • Privacy Policy
  • Cookie Policy
RunComfy
Copyright 2026 RunComfy. All Rights Reserved.

RunComfy is the premier ComfyUI platform, offering ComfyUI online environment and services, along with ComfyUI workflows featuring stunning visuals. RunComfy also provides AI Models, enabling artists to harness the latest AI tools to create incredible art.

Examples Of Seedance 2.5 Reference to Video

Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...
Video thumbnail
Loading...