SAFE REINFORCEMENT LEARNING MODEL SERVICE

Patent №

US 11,640,345

Granted

2023-05-02

Filed 2020

Owner

AMAZON TECHNOLOGIES, INC.

AI components

5

ml · vision · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16875791

Systems and methods are described for training a machine learning model to make a series of sequential decisions, in which the results of previous decisions are known prior to the next decision in the sequence being made. A safe reinforcement learning model estimates the results of choosing various options for a first decision in the sequence, and further estimates the amount of information that will be gained by choosing each of the options. The estimated information gain associated with each option is then used to forecast how the remaining decisions in the sequence would be improved by using the gained information to improve the prediction model and make better decisions. The safe reinforcement learning model further incorporates decision constraints provided by subject matter experts, which may set requirements for the selection such as a minimum required result and allow the safe reinforcement learning model to explore options within those constraints.

Machine learningVisionKnowledge representationPlanningAI hardwareG06F 11/30G06F 18/2113G06F 18/214G06N 5/01G06N 5/04G06N 7/01G06N 20/00G06F 11/3409+1 more

AI classification

Machine learning1.00
Planning1.00
AI hardware1.00
Knowledge representation1.00
Vision0.98
Natural language0.32
Evolutionary computation0.13
Speech0.01

Ownership

AMAZON TECHNOLOGIES, INC.

assignment · 550840984

Assignors

GHOSH, NEAL KISHORE, LABER, ERIC BENJAMIN

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC