SYSTEM AND METHOD FOR EXTRACTING TEXT CAPTIONS FROM VIDEO AND GENERATING VIDEO SUMMARIES

Patent №

US 8,488,682

Granted

2013-07-16

Filed 2007

Owner

THE TRUSTEES OF COLUMBIA UNIVERSITY IN THE CITY OF NEW YORK

Lab

AI components

4

ml · nlp · vision · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

11960424

Caption boxes which are embedded in video content can be located and the text within the caption boxes decoded. Real time processing is enhanced by locating caption box regions in the compressed video domain and performing pixel based processing operations within the region of the video frame in which a caption box is located. The captions boxes are further refined by identifying word regions within the caption boxes and then applying character and word recognition processing to the identified word regions. Domain based models are used to improve text recognition results. The extracted caption box text can be used to detect events of interest in the video content and a semantic model applied to extract a segment of video of the event of interest.

Machine learningNatural languageVisionSpeechG11B 27/031G06F 16/739G06F 16/7844G06F 16/7857G06F 16/786G06T 7/75G06V 20/635G11B 27/034+7 more

AI classification

Natural language1.00
Vision1.00
Machine learning1.00
Speech0.99
Knowledge representation0.07
AI hardware0.00
Planning0.00
Evolutionary computation0.00

Ownership

THE TRUSTEES OF COLUMBIA UNIVERSITY IN THE CITY OF NEW YORK

assignment · 214020226

From the same owner

© 2026 NYSGPT2525 LLC