Skip to main content
The WanSoundImageToVideo node prepares conditioning and an empty latent video tensor for Wan sound-to-video generation. It can optionally incorporate audio encoding, a reference image, a control video, and a motion reference to guide the generated video.

Inputs

Note: All optional inputs can be used independently or together. The node modifies the provided positive and negative conditioning based on which optional inputs are connected.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): b1148cd00d8999dd6842e3c2fb13655fda8f20d5befed975a6d1652688b2807c