minimax-h3-open/reference-to-video
Open-weights reference-to-video from image, video, and audio cues.
Open-weights reference-to-video from image, video, and audio cues.
Open-weights image-to-video with optional last frame and native stereo audio.
Open-weights text-to-video with 480p/768p and native stereo audio.
FLUX 3 Draft Extend: Fast low-cost clip continuation drafts at 720p
FLUX 3 Draft Keyframes: Fast multi-keyframe video previews at 720p
FLUX 3 Draft FLF: Fast start-end frame video previews at 720p
FLUX 3 Draft I2V: Fast still-to-video preview drafts at 720p
FLUX 3 Draft: Fast, low-cost text-to-video previews at 720p
Seedance 2.5 FLF2V 480p: First-last frame draft transitions at lower cost
Seedance 2.5 Reference 480p: Multi-reference draft video at lower cost
Seedance 2.5 480p: Fast, low-cost still-to-video animation drafts
Seedance 2.5 480p: Fast, low-cost text-to-video drafts from prompts
Lengthen existing clips beyond the final frame with scene-consistent motion
Generate video from multi-keyframe stills with optional audio
Generate video between start and end frames with optional audio
Animate a start image into video with optional native audio
Bridge start and end stills into smooth cinematic video
Generate 768p or 2K video from image, video, and audio references
Animate a first-frame image into 768p or 2K video up to 15 seconds
MiniMax H3: 768p/2K text-to-video with native stereo audio
FLUX 3 Video turns text prompts into cinematic clips with native audio
FLUX 3 Image is a multimodal image model with reference guidance and readable text
Edit a source video from a text instruction while keeping scene coherence.
Turn reference images and a prompt into short video with synced audio.
Animate a still image into a short video with synchronized audio.
Generate cinematic videos with synchronized audio from a text prompt.
Seedance 2.5 Reference to Video: Turn reference images, videos, and audio into cinematic AI video
Seedance 2.5: Animate a still image into cinematic AI video
Seedance 2.5: Cinematic AI video with stronger consistency and longer clips
Reference-guided image editing with layout, layer separation, and multilingual text
Text-to-image and image editing with layout and layer control
Fast, low-cost prompt-based image editing at a fixed 1K resolution.
Fast, low-cost text-to-image generation at a fixed 1K resolution.
Generates natural speech and audio from text, reference audio, or an image
Reshape a source clip from a text prompt with native audio.
Animate a start image into a cinematic clip with native audio.
Fast, low-cost multi-shot AI video model with native audio and references.
Turn reference images into smooth 720P or 1080P video with one prompt.
Animate a still photo into smooth 720P or 1080P video from one prompt.
Multimodal AI video model with native audio for text, image, and reference inputs.
Generate posters, logos, and typography-rich images from text prompts.
Prompt-driven video editing at $0.126 per second of output.
Cinematic 4K image-to-video at $0.42 per second of output.
Cinematic 4K reference-to-video at $0.42 per second of output.
Cinematic 4K text-to-video at $0.42 per second of output.
Pro-tier image animation: 3-15s cinematic clips from $0.112 per second.
Pro-tier reference-driven 3-15s video generation from $0.112 per second.
Cinematic Pro-tier text-to-video at $0.112 per second of output.
Prompt-driven Pro-tier video editing at $0.168 per second.
Image-to-video 3-15s clips at $0.084 per second.
RunComfy is the premier ComfyUI platform, offering ComfyUI online environment and services, along with ComfyUI workflows featuring stunning visuals. RunComfy also provides AI Models, enabling artists to harness the latest AI tools to create incredible art.
