Save 4 hours! We auto-setup your workflow! Free!

Drop your workflow.json — we handle every dependency, custom node, and model. Just open the link and run.

Auto-Setup Workflow Json (Free) Now!
ComfyUI > Nodes > ComfyUI-CustomNodeKit > Wan SCAIL To Video (Multi Ref)

ComfyUI Node: Wan SCAIL To Video (Multi Ref)

Class Name

WanSCAILToVideoMultiRef

Category
model/conditioning/video_models
Author
user2318 (Account age: 2574days)
Extension
ComfyUI-CustomNodeKit
Latest Updated
2026-07-16
Github Stars
0.05K

How to Install ComfyUI-CustomNodeKit

Install this extension via the ComfyUI Manager by searching for ComfyUI-CustomNodeKit
  • 1. Click the Manager button in the main menu
  • 2. Select Custom Nodes Manager button
  • 3. Enter ComfyUI-CustomNodeKit in the search bar
After installation, click the Restart button to restart ComfyUI. Then, manually refresh your browser to clear the cache and access the updated list of nodes.

Visit ComfyUI Online for ready-to-use ComfyUI environment

  • Free trial available
  • 16GB VRAM to 80GB VRAM GPU machines
  • 400+ preloaded models/nodes
  • Freedom to upload custom models/nodes
  • 200+ ready-to-run workflows
  • 100% private workspace with up to 200GB storage
  • Dedicated Support

Run ComfyUI Online

Wan SCAIL To Video (Multi Ref) Description

Facilitates conditioning of video models with multiple reference images and masks for enhanced video generation.

Wan SCAIL To Video (Multi Ref):

The WanSCAILToVideoMultiRef node is designed to facilitate the conditioning of video models using multiple reference images and masks. This node is part of a broader system that leverages SCAIL (Spatially Conditioned Adversarial Image Learning) techniques to enhance video generation by incorporating pose and mask data. It allows for the integration of multiple reference images, which can be used to guide the generation of video content, ensuring that the output adheres to specific visual styles or characteristics defined by these references. The node is particularly useful for applications where maintaining consistency across frames is crucial, such as in animation or video synthesis tasks. By supporting multi-reference inputs, it provides flexibility and control over the generated video content, making it a valuable tool for AI artists looking to create complex and visually coherent video sequences.

Wan SCAIL To Video (Multi Ref) Input Parameters:

pose_video

The pose_video parameter is a tensor representing the video frames that contain pose information. This input is crucial as it provides the foundational data upon which the video generation is conditioned. The pose information helps in aligning the generated content with the desired movements or actions depicted in the video. There are no specific minimum or maximum values, but the quality and resolution of the input video can significantly impact the results.

pose_video_mask

The pose_video_mask parameter is an optional input that provides a mask for the pose video. This mask is used to define areas of interest or focus within the video frames, allowing the node to apply conditioning more precisely. The mask can help in isolating specific regions, such as a character or object, ensuring that the conditioning effects are applied only where needed. The mask should match the dimensions of the pose video for optimal results.

positive

The positive parameter is a conditioning input that allows you to set specific values or conditions for the video generation process. It is used to guide the model towards desired outcomes by emphasizing certain features or characteristics. This parameter can be adjusted to influence the strength and nature of the conditioning applied to the video.

negative

The negative parameter works in conjunction with the positive parameter to provide a balanced conditioning approach. While the positive input emphasizes certain features, the negative input can be used to de-emphasize or suppress unwanted characteristics. This dual conditioning approach helps in achieving a more refined and controlled video output.

reference_image_mask

The reference_image_mask is a tensor that provides a mask for the reference images used in the conditioning process. This mask helps in defining which parts of the reference images should influence the video generation. By using this mask, you can ensure that only relevant features from the reference images are considered, enhancing the coherence and quality of the generated video.

Wan SCAIL To Video (Multi Ref) Output Parameters:

positive

The positive output parameter contains the conditioned video data that has been influenced by the positive conditioning inputs. This output reflects the desired characteristics and features emphasized during the conditioning process, providing a video sequence that aligns with the specified artistic or stylistic goals.

negative

The negative output parameter provides the conditioned video data that has been influenced by the negative conditioning inputs. This output helps in understanding the effects of de-emphasizing certain features, offering a contrast to the positive output and allowing for a more comprehensive evaluation of the conditioning process.

out_latent

The out_latent output parameter is a dictionary containing the latent representations of the video frames. This includes the samples key, which holds the latent data, and optionally a noise_mask key if noise masking was applied. The latent representations are crucial for understanding the underlying structure and features of the generated video, providing insights into the model's internal processing.

Wan SCAIL To Video (Multi Ref) Usage Tips:

  • Ensure that the pose video and pose video mask are of high quality and resolution to achieve the best results in video generation.
  • Experiment with different combinations of positive and negative conditioning inputs to fine-tune the output and achieve the desired artistic effects.
  • Utilize the reference image mask to focus the conditioning effects on specific areas of the reference images, enhancing the coherence and relevance of the generated video content.

Wan SCAIL To Video (Multi Ref) Common Errors and Solutions:

"Shape mismatch between pose video and mask"

  • Explanation: This error occurs when the dimensions of the pose video and the pose video mask do not match, leading to issues in applying the mask correctly.
  • Solution: Ensure that the pose video and pose video mask have the same dimensions. Resize or adjust the mask to match the video frames before inputting them into the node.

"Invalid reference image mask"

  • Explanation: This error indicates that the reference image mask is not properly formatted or does not align with the expected input structure.
  • Solution: Verify that the reference image mask is correctly formatted and matches the dimensions of the reference images. Adjust the mask as needed to ensure compatibility with the node's requirements.

Wan SCAIL To Video (Multi Ref) Related Nodes

Go back to the extension to check out more related nodes.
ComfyUI-CustomNodeKit
RunComfy
Copyright 2025 RunComfy. All Rights Reserved.

RunComfy is the premier ComfyUI platform, offering ComfyUI online environment and services, along with ComfyUI workflows featuring stunning visuals. RunComfy also provides AI Models, enabling artists to harness the latest AI tools to create incredible art.

Wan SCAIL To Video (Multi Ref)