Semantic Parsing of Objects in Video

Patent №

US 8,774,522

Granted

2014-07-08

Filed 2013

Owner

INTERNATIONAL BUSINESS MACHINES CORPORATION

AI components

3

nlp · vision · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

13948325

Methods, systems, and computer program products for parsing objects in a video are provided herein. A method includes producing a plurality of versions of an image of an object, wherein each version has a different resolution of said image of said object, and computing an appearance score at each of a plurality of regions on the lowest resolution version for at least one semantic attribute for said object. Such a method also includes analyzing one or more other versions to compute a resolution context score for each of the plurality of regions in the lowest resolution version, wherein said resolution context score denotes an extent to which finer spatial structure exists in the one or more others versions than in the lowest resolution version, and determining a configuration of the at least one semantic attribute in the lowest resolution version based on the appearance score and the resolution context score.

Natural languageVisionAI hardwareG06V 20/10G06F 18/22G06V 10/426G06V 20/41G06V 20/70G06V 30/2504G06V 40/103

AI classification

Vision1.00
Natural language1.00
AI hardware0.88
Knowledge representation0.40
Evolutionary computation0.01
Machine learning0.00
Planning0.00
Speech0.00

Ownership

INTERNATIONAL BUSINESS MACHINES CORPORATION

assignment · 512530046

Assignors

BROWN, LISA MARIE, FERIS, ROGERIO SCHMIDT, HAMPAPUR, ARUN, VAQUERO, DANIEL ANDRE

On an employer assignment, the assignors are typically the inventors.

From the same owner

© 2026 NYSGPT2525 LLC