Patent №
US 10,445,641
Granted
2019-10-15
Filed 2016
Owner
GOOGLE INC.
AI components
5
ml · vision · kr · planning · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
15016173
Methods, systems, and apparatus, including computer programs encoded on computer storage media, for distributed training of reinforcement learning systems. One of the methods includes receiving, by a learner, current values of the parameters of the Q network from a parameter server, wherein each learner maintains a respective learner Q network replica and a respective target Q network replica; updating, by the learner, the parameters of the learner Q network replica maintained by the learner using the current values; selecting, by the learner, an experience tuple from a respective replay memory; computing, by the learner, a gradient from the experience tuple using the learner Q network replica maintained by the learner and the target Q network replica maintained by the learner; and providing, by the learner, the computed gradient to the parameter server.
AI classification
Ownership
GOOGLE INC.
assignment · 377300402
Assignors
SRINIVASAN, PRAVEEN DEEPAK, FEARON, RORY, ALCICEK, CAGDAS, NAIR, ARUN SARATH, BLACKWELL, SAMUEL, PANNEERSHELVAM, VEDAVYAS, DE MARIA, ALESSANDRO, MNIH, VOLODYMYR, KAVUKCUOGLU, KORAY, SILVER, DAVID, SULEYMAN, MUSTAFA
On an employer assignment, the assignors are typically the inventors.