METHOD AND APPARATUS FOR GENERATING A CONDENSED VERSION OF A VIDEO SEQUENCE INCLUDING DESIRED AFFORDANCES

Patent №

US 6,560,281

Granted

2003-05-06

Filed 1998

Owner

XEROX CORPORATION

Lab

AI components

4

nlp · vision · kr · planning

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

09028548

A method and apparatus analyzes and annotates a technical talk typically illustrated with overhead slides, wherein the slides are recorded in a video sequence. The video sequence is condensed and digested into key video frames adaptable for annotation to time and audio sequence. The system comprises a recorder for recording a technical talk as a sequential set of video image frames. A stabilizing processor segregates the video image frames into a plurality of associated subsets each corresponding to a distinct slide displayed at the talk and for median filtering of the subsets for generating a key frame representative of each of the subsets. A comparator compares the key frame with the associated subsets to identify differences between the key frame and the associates subset which comprise nuisances and affordances. A gesture recognizer locates, tracks and recognizes gestures occurring in the subset as gesture affordances and identifies a gesture video frame representative of the gesture affordance. An integrator compiles the key frames and gesture video frames as a digest of the video image frames which can also be annotated with the time and audio sequence.

AI classification

Vision1.00
Knowledge representation0.97
Planning0.96
Natural language0.95
AI hardware0.38
Speech0.08
Machine learning0.04
Evolutionary computation0.00

Ownership

XEROX CORPORATION

assignment · 91020442

Assignors

BLACK, MICHAEL J., JU, XUAN, MINNEMAN, SCOTT, KIMBER, DONALD G.

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC