Inputs
Parameter Constraints:
- When
pose_videois provided, the output length will be adjusted to match the pose video duration if thetrim_to_pose_videologic is active (currently set toFalsein the source code) face_videois automatically resized to 512x512 resolution and normalized to a range of -1.0 to 1.0 when processedcontinue_motionframes are limited by thecontinue_motion_max_framesparameter; only the lastcontinue_motion_max_framesframes from the input are used- Input videos (
face_video,pose_video,background_video,character_mask) are offset byvideo_frame_offsetbefore processing; if the offset exceeds the video length, the input is ignored - If
character_maskcontains only one frame, it will be repeated across all frames - When
clip_vision_outputis provided, it’s applied to both positive and negative conditioning - If
reference_imageis not provided, a black image (all zeros) is used as the default reference - If
continue_motionis not provided, the initial frames are filled with gray (0.5 intensity) noise
Outputs
This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub
Source fingerprint (SHA-256):
2ec2afbc57f58a5b7ce0ecc3730618633d435439ce2d650b18be531c1edddff0