Skip to main content
The LTXVAddGuide node encodes input images or videos through a VAE encoder and adds them as guide keyframes to a latent video sequence. It updates both positive and negative conditioning and returns the modified latent, with options to set the starting frame, conditioning strength, an optional attention mask, and optional IC-LoRA parameters.

Inputs

Note: The input image/video must have a frame count following the 8*n + 1 pattern (e.g., 1, 9, 17, 25 frames). If the input exceeds this pattern, it will be automatically cropped to the nearest valid frame count. Note on iclora_parameters: When using IC-LoRA parameters with a reference_downscale_factor greater than 1, the latent spatial dimensions (width and height) must be divisible by that factor. The node will raise an error if this condition is not met. Note: The encoded guide frames must fit within the latent sequence at the selected frame position. If the conditioned frames exceed the length of the latent sequence, the node raises an error. Note: Adding a guide to a latent that combines audio and video channels is not supported and will raise an error.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 031bc9030dafed85b5ff1cbceae36234e9d5f77f7f4b040267067ecd16a27929