Patent №
US 12,670,722
Granted
2026-06-30
Filed 2022
Owner
CARNEGIE MELLON UNIVERSITY
AI components
0
Assignment
None on record
Dataset
AIPD
Application
18077974
A method of compositional feature representation learning for video understanding is described. The method includes individually processing a sequence of video frames received as an input of a feature map network to generate a plurality of feature maps. The method also includes binding the plurality of feature maps to a fixed set of slot variables using an attention model according to a motion segmentation signal. The method further includes combining slot states corresponding to the fixed set of slot variables into a combined feature map. The method also includes decoding the combined feature map to form a reconstructed sequence of video frames, in which objects discovered in the reconstructed sequence of video frames are identified.
Ownership
CARNEGIE MELLON UNIVERSITY