EXTRACTION OF TEXT ELEMENTS FROM VIDEO CONTENT

Patent №

US 8,340,498

Granted

2012-12-25

Filed 2009

Owner

AMAZON TECHNOLOGIES, INC.

AI components

5

nlp · vision · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

12364944

Video content comprising a plurality of frames containing textual and non-textual elements is processed. A portion of the plurality of frames is selected for analysis to identify textual elements in the frames corresponding to pre-defined textual elements. The identified textual elements are stored along with their location within the video content. In some embodiments, each of a subset of frames included in the portion is analyzed until the pre-defined textual element is identified in a start frame. A plurality of successive frames subsequent to the start frame is analyzed to identify pre-defined textual elements in the frame. Analyzing the frames includes filtering the frames to remove non-textual elements and increase the visibility of the textual elements contained therein. A confidence rating is calculated for the identified textual elements according to some embodiments.

AI classification

Vision1.00
Natural language1.00
Planning1.00
Knowledge representation0.85
AI hardware0.64
Machine learning0.03
Evolutionary computation0.01
Speech0.00

Ownership

AMAZON TECHNOLOGIES, INC.

assignment · 282420142

Assignors

GILL, SUNBIR, TALREJA, KAMLESH T., SCHWABLAND, PETER A.

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC