base_shift and max_shift across a token range, then patches the input model with a specialized sampling configuration. If a latent is provided, its dimensions determine the token count; otherwise 4096 tokens are used.
Inputs
The shift value is calculated by interpolating between
base_shift at 1024 tokens and max_shift at 4096 tokens. When latent is supplied, the token count is the product of all dimensions after the first two in the latent samples (the spatial/temporal dimensions). If no latent is provided, the token count defaults to 4096.
Outputs
This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub
Source fingerprint (SHA-256):
aba596c5478e9d6ee821eec1eca15506935bcc765a368087ccc442fc2ed6671b