METHOD AND APPARATUS FOR GENERATING A CONDENSED VERSION OF A VIDEO SEQUENCE INCLUDING DESIRED AFFORDANCES
Patent №
US 6,560,281
Granted
2003-05-06
Filed 1998
Owner
XEROX CORPORATION
Lab
—
AI components
4
nlp · vision · kr · planning
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
09028548
A method and apparatus analyzes and annotates a technical talk typically illustrated with overhead slides, wherein the slides are recorded in a video sequence. The video sequence is condensed and digested into key video frames adaptable for annotation to time and audio sequence. The system comprises a recorder for recording a technical talk as a sequential set of video image frames. A stabilizing processor segregates the video image frames into a plurality of associated subsets each corresponding to a distinct slide displayed at the talk and for median filtering of the subsets for generating a key frame representative of each of the subsets. A comparator compares the key frame with the associated subsets to identify differences between the key frame and the associates subset which comprise nuisances and affordances. A gesture recognizer locates, tracks and recognizes gestures occurring in the subset as gesture affordances and identifies a gesture video frame representative of the gesture affordance. An integrator compiles the key frames and gesture video frames as a digest of the video image frames which can also be annotated with the time and audio sequence.
AI classification
Ownership
XEROX CORPORATION
assignment · 91020442
Assignors
BLACK, MICHAEL J., JU, XUAN, MINNEMAN, SCOTT, KIMBER, DONALD G.
On an employer assignment, the assignors are typically the inventors.