SELF-SUPERVISED COMPOSITIONAL FEATURE REPRESENTATION FOR VIDEO UNDERSTANDING

Patent №

US 12,670,722

Granted

2026-06-30

Filed 2022

Owner

CARNEGIE MELLON UNIVERSITY

AI components

0

Assignment

None on record

Dataset

AIPD

Application

18077974

A method of compositional feature representation learning for video understanding is described. The method includes individually processing a sequence of video frames received as an input of a feature map network to generate a plurality of feature maps. The method also includes binding the plurality of feature maps to a fixed set of slot variables using an attention model according to a motion segmentation signal. The method further includes combining slot states corresponding to the fixed set of slot variables into a combined feature map. The method also includes decoding the combined feature map to form a reconstructed sequence of video frames, in which objects discovered in the reconstructed sequence of video frames are identified.

G06V 20/58G06V 10/82G06T 7/215G06T 7/246G06V 10/454G06V 10/7715G06V 20/46G06V 10/7753+4 more

Ownership

CARNEGIE MELLON UNIVERSITY

© 2026 NYSGPT2525 LLC