Patent №
US 11,640,345
Granted
2023-05-02
Filed 2020
Owner
AMAZON TECHNOLOGIES, INC.
Lab
AI components
5
ml · vision · kr · planning · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
16875791
Systems and methods are described for training a machine learning model to make a series of sequential decisions, in which the results of previous decisions are known prior to the next decision in the sequence being made. A safe reinforcement learning model estimates the results of choosing various options for a first decision in the sequence, and further estimates the amount of information that will be gained by choosing each of the options. The estimated information gain associated with each option is then used to forecast how the remaining decisions in the sequence would be improved by using the gained information to improve the prediction model and make better decisions. The safe reinforcement learning model further incorporates decision constraints provided by subject matter experts, which may set requirements for the selection such as a minimum required result and allow the safe reinforcement learning model to explore options within those constraints.
AI classification
Ownership
AMAZON TECHNOLOGIES, INC.
assignment · 550840984
Assignors
GHOSH, NEAL KISHORE, LABER, ERIC BENJAMIN
On an employer assignment, the assignors are typically the inventors.