MULTI-AGENT REINFORCEMENT LEARNING WITH MATCHMAKING POLICIES

Patent №

US 11,627,165

Granted

2023-04-11

Filed 2020

Owner

DEEPMIND TECHNOLOGIES LIMITED

AI components

4

ml · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16752496

Methods, systems, and apparatus, including computer programs encoded on a computer storage medium, for training a policy neural network having a plurality of policy parameters and used to select actions to be performed by an agent to control the agent to perform a particular task while interacting with one or more other agents in an environment. In one aspect, the method includes: maintaining data specifying a pool of candidate action selection policies; maintaining data specifying respective matchmaking policy; and training the policy neural network using a reinforcement learning technique to update the policy parameters. The policy parameters define policies to be used in controlling the agent to perform the particular task.

Machine learningKnowledge representationPlanningAI hardwareG06N 3/08G06F 18/214G06N 3/006G06N 3/045G06N 3/09G06N 3/092G06N 3/0985H04L 63/205

AI classification

Planning1.00
Machine learning1.00
Knowledge representation1.00
AI hardware0.96
Speech0.42
Evolutionary computation0.12
Vision0.01
Natural language0.00

Ownership

DEEPMIND TECHNOLOGIES LIMITED

assignment · 517700438

Assignors

SILVER, DAVID, VINYALS, ORIOL, JADERBERG, MAXWELL ELLIOT

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC