Skip to main content
Track objects across video frames using SAM3’s memory-based tracker. The node processes a sequence of video frames and maintains object identities across frames, using either initial masks or text prompts to define what to track, and can detect new objects along the way using text conditioning.

Inputs

Note: Either initial_mask or conditioning must be provided. If both are omitted, the node raises an error.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): ef584628b334997a001a857a7deffb7eda34db8fa50e3d734a07b5e92566d48d