ONLINE AGENT USING REINFORCEMENT LEARNING TO PLAN AN OPEN SPACE TRAJECTORY FOR AUTONOMOUS VEHICLES

Patent №

US 11,467,591

Granted

2022-10-11

Filed 2019

Owner

BAIDU USA LLC

Lab

AI components

5

ml · vision · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16413332

In one embodiment, a system uses an actor-critic reinforcement learning model to generate a trajectory for an autonomous driving vehicle (ADV) in an open space. The system perceives an environment surrounding an ADV. The system applies a RL algorithm to an initial state of a planning trajectory based on the perceived environment to determine a plurality of controls for the ADV to advance to a plurality of trajectory states based on map and vehicle control information for the ADV. The system determines a reward prediction by the RL algorithm for each of the plurality of controls in view of a target destination state. The system generates a first trajectory from the trajectory states by maximizing the reward predictions to control the ADV autonomously according to the first trajectory.

Machine learningVisionKnowledge representationPlanningAI hardwareG05D 1/0221G06N 3/006B60W 60/0011G05D 1/0217G06N 3/045G06N 3/0464G06N 3/088G06N 3/092+1 more

AI classification

Planning1.00
Machine learning1.00
AI hardware1.00
Vision0.86
Knowledge representation0.62
Natural language0.01
Speech0.00
Evolutionary computation0.00

Ownership

BAIDU USA LLC

assignment · 496590412

Assignors

HE, RUNXIN, ZHOU, JINYUN, LUO, QI, SONG, SHIYU, MIAO, JINGHAO, HU, JIANGTAO, WANG, YU, XU, JIAXUAN, JIANG, SHU

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC