Skip to main content
The WanDancerVideo node prepares conditioning data and an empty latent tensor for video generation with the WanDancer model. It combines positive and negative conditioning with optional inputs like a starting image, mask, CLIP vision embeddings, and audio features to control the generated video.

Inputs

Note on Parameter Constraints:
  • The start_image and mask inputs are optional but can be used together. When start_image is provided, it is encoded and concatenated with the latent. If mask is also provided, it controls which parts of the start image are kept (white) and which are regenerated (black). If mask is not provided, the entire start image area is used as a conditioning guide.
  • The clip_vision_output and clip_vision_output_ref inputs are optional and can be used together to provide visual context for the first frame and a reference image.
  • The audio_encoder_output input is optional and provides audio features for audio-conditional generation.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 0a75b24c8e5c164d81b08eb438862d94d4409ece8dc22c126979347e2350c828