DEEP REINFORCEMENT LEARNING FRAMEWORK FOR CHARACTERIZING VIDEO CONTENT

Patent №

US 10,885,341

Granted

2021-01-05

Filed 2018

Owner

SONY INTERACTIVE ENTERTAINMENT INC.

Lab

AI components

7

ml · nlp · vision · speech · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16171018

Methods and systems for performing sequence level prediction of a video scene are described. Video information in a video scene is represented as a sequence of features depicted each frame. An environment state for each time step t corresponding to each frame is represented by the video information for time step t and predicted affective information from a previous time step t−1. An action A(t) as taken with an agent controlled by a machine learning algorithm for the frame at step t, wherein an output of the action A(t) represents affective label prediction for the frame at the time step t. A pool of predicted actions is transformed to a predicted affective history at a next time step t+1. The predictive affective history is included as part of the environment state for the next time step t+1. A reward R is generated on predicted actions up to the current time step t, by comparing them against corresponding annotated movie scene affective labels.

Machine learningNatural languageVisionSpeechKnowledge representationPlanningAI hardwareG06N 3/045G06N 3/08G06F 18/217G06N 3/006G06N 3/0464G06N 3/0895G06N 3/092G06N 20/00+3 more

AI classification

Machine learning1.00
Vision1.00
Knowledge representation0.99
Planning0.93
Natural language0.83
Speech0.82
AI hardware0.73
Evolutionary computation0.04

Ownership

SONY INTERACTIVE ENTERTAINMENT INC.

assignment · 483620447

Assignors

CHEN, RUXIN, KUMAR, NAVEEN, LI, HAOQI

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC