EYE GAZE DRIVEN SPATIO-TEMPORAL ACTION LOCALIZATION

Patent №

US 9,514,363

Granted

2016-12-06

Filed 2014

Owner

DISNEY ENTERPRISES, INC.

Lab

AI components

5

ml · nlp · vision · kr · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

14247845

The disclosure provides an approach for detecting and localizing action in video. In one embodiment, an action detection application receives training video sequences and associated eye gaze fixation data collected from a sample of human viewers. Using the training video sequences and eye gaze data, the action detection application learns a model which includes a latent regions potential term that measures the compatibility of latent spatio-temporal regions with the model, as well as a context potential term that accounts for contextual information that is not directly produced by the appearance and motion of the actor. The action detection application may train this model in, e.g., the latent structural SVM framework by minimizing a cost function which encodes the cost of an incorrect action label prediction and a mislocalization of the eye gaze. During training and thereafter, inferences using the model may be made using an efficient dynamic programming algorithm.

Machine learningNatural languageVisionKnowledge representationAI hardwareG06V 10/7784G06F 18/2178G06F 18/24G06V 10/25G06V 40/193G06V 40/20

AI classification

Vision1.00
Machine learning1.00
Natural language1.00
AI hardware0.98
Knowledge representation0.79
Planning0.00
Evolutionary computation0.00
Speech0.00

Ownership

DISNEY ENTERPRISES, INC.

assignment · 326280992

Assignors

SHAPOVALOVA, NATALIYA, SIGAL, LEONID, RAPTIS, MICHAIL

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC