REINFORCEMENT LEARNING THROUGH A DOUBLE ACTOR CRITIC ALGORITHM

Patent №

US 11,816,591

Granted

2023-11-14

Filed 2020

Owner

SONY CORPORATION OF AMERICA

+1 more

Lab

AI components

6

ml · vision · speech · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16800463

The Double Actor Critic (DAC) reinforcement-learning algorithm affords stable policy improvement and aggressive neural-net optimization without catastrophic overfitting of the policy. DAC trains models using an arbitrary history of data in both offline and online learning and can be used to smoothly improve on an existing policy learned or defined by some other means. Finally, DAC can optimize reinforcement learning problems with discrete and continuous action spaces.

Machine learningVisionSpeechKnowledge representationPlanningAI hardwareG06N 3/084G06N 7/01G06N 3/006G06N 3/045G06N 3/0499G06N 3/092G06N 20/00

AI classification

Machine learning1.00
AI hardware1.00
Knowledge representation1.00
Planning0.99
Vision0.90
Speech0.64
Natural language0.01
Evolutionary computation0.00

Ownership

SONY CORPORATION OF AMERICA

assignment · 519210439

SONY CORPORATION

assignment · 519210439

Assignors

COGITAI, INC.

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC