SYSTEM AND METHOD FOR EXTRACTING TEXT CAPTIONS FROM VIDEO AND GENERATING VIDEO SUMMARIES
Patent №
US 8,488,682
Granted
2013-07-16
Filed 2007
Owner
THE TRUSTEES OF COLUMBIA UNIVERSITY IN THE CITY OF NEW YORK
Lab
—
AI components
4
ml · nlp · vision · speech
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
11960424
Caption boxes which are embedded in video content can be located and the text within the caption boxes decoded. Real time processing is enhanced by locating caption box regions in the compressed video domain and performing pixel based processing operations within the region of the video frame in which a caption box is located. The captions boxes are further refined by identifying word regions within the caption boxes and then applying character and word recognition processing to the identified word regions. Domain based models are used to improve text recognition results. The extracted caption box text can be used to detect events of interest in the video content and a semantic model applied to extract a segment of video of the event of interest.
AI classification
Ownership
THE TRUSTEES OF COLUMBIA UNIVERSITY IN THE CITY OF NEW YORK
assignment · 214020226