Skip to main content
This node analyzes an audio clip and turns it into a set of features that can guide a video generation model. It estimates tempo and beats, extracts mel-spectrogram, MFCC, chroma, and onset features, then packages them together with a calculated frame rate for synchronization.

Inputs

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): ce27a3bdea2d9e3cf8875c24236a2a0a1429e9bc13a58581e372fb669d2c0018