ONLINE AGENT USING REINFORCEMENT LEARNING TO PLAN AN OPEN SPACE TRAJECTORY FOR AUTONOMOUS VEHICLES
Patent №
US 11,467,591
Granted
2022-10-11
Filed 2019
Owner
BAIDU USA LLC
Lab
—
AI components
5
ml · vision · kr · planning · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
16413332
In one embodiment, a system uses an actor-critic reinforcement learning model to generate a trajectory for an autonomous driving vehicle (ADV) in an open space. The system perceives an environment surrounding an ADV. The system applies a RL algorithm to an initial state of a planning trajectory based on the perceived environment to determine a plurality of controls for the ADV to advance to a plurality of trajectory states based on map and vehicle control information for the ADV. The system determines a reward prediction by the RL algorithm for each of the plurality of controls in view of a target destination state. The system generates a first trajectory from the trajectory states by maximizing the reward predictions to control the ADV autonomously according to the first trajectory.
AI classification
Ownership
BAIDU USA LLC
assignment · 496590412
Assignors
HE, RUNXIN, ZHOU, JINYUN, LUO, QI, SONG, SHIYU, MIAO, JINGHAO, HU, JIANGTAO, WANG, YU, XU, JIAXUAN, JIANG, SHU
On an employer assignment, the assignors are typically the inventors.