Patent №
US 11,543,789
Granted
2023-01-03
Filed 2020
Owner
FUJITSU LIMITED
Lab
—
AI components
4
ml · planning · evo · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
16797515
A reinforcement learning method executed by a computer includes calculating a degree of risk for a state of a controlled object at a current time point with respect to a constraint condition related to the state of the controlled object, the degree of risk being calculated based on a predicted value of the state of the controlled object at a future time point, the predicted value being obtained from model information defining a relationship between the state of the controlled object and a control input to the controlled object; and determining the control input to the controlled object at the current time point, from a range defined according to the calculated degree of risk so that the range becomes narrower as the calculated degree of risk increases.
AI classification
Ownership
FUJITSU LIMITED
assignment · 518880475
Assignors
OKAWA, YOSHIHIRO, SASAKI, TOMOTAKE, IWANE, HIDENAO, YANAMI, HITOSHI
On an employer assignment, the assignors are typically the inventors.