Patent №
US 11,657,266
Granted
2023-05-23
Filed 2018
Owner
HONDA MOTOR CO., LTD.
Lab
—
AI components
5
ml · vision · kr · planning · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
16193291
According to one aspect, cooperative multi-goal, multi-agent, multi-stage (CM3) reinforcement learning may include training a first agent using a first policy gradient and a first critic using a first loss function to learn goals in a single-agent environment using a Markov decision process, training a number of agents based on the first policy gradient and a second policy gradient and a second critic based on the first loss function and a second loss function to learn cooperation between the agents in a multi-agent environment using a Markov game to instantiate a second agent neural network, each of the agents instantiated with the first agent neural network in a pre-trained fashion, and generating a CM3 network policy based on the first agent neural network and the second agent neural network. The CM3 network policy may be implemented in a CM3 based autonomous vehicle to facilitate autonomous driving.
AI classification
Ownership
HONDA MOTOR CO., LTD.
assignment · 475250974
Assignors
NAKHAEI SARVEDANI, ALIREZA, ISELE, DAVID FRANCIS
On an employer assignment, the assignors are typically the inventors.