RECORDING MEDIUM, REINFORCEMENT LEARNING METHOD, AND REINFORCEMENT LEARNING APPARATUS

Patent №

US 11,645,574

Granted

2023-05-09

Filed 2018

Owner

FUJITSU LIMITED

+1 more

Lab

AI components

4

ml · planning · evo · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16130482

A non-transitory, computer-readable recording medium stores therein a reinforcement learning program that uses a value function and causes a computer to execute a process comprising: estimating first coefficients of the value function represented in a quadratic form of inputs at times in the past than a present time and outputs at the present time and the times in the past, the first coefficients being estimated based on inputs at the times in the past, the outputs at the present time and the times in the past, and costs or rewards that corresponds to the inputs at the times in the past; and determining second coefficients that defines a control law, based on the value function that uses the estimated first coefficients and determining input values at times after estimation of the first coefficients.

Machine learningPlanningEvolutionary computationAI hardwareG06N 20/00G06N 3/006H04L 41/0816H04L 41/16H04L 43/08

AI classification

Machine learning1.00
Evolutionary computation1.00
Planning1.00
AI hardware1.00
Knowledge representation0.13
Vision0.01
Natural language0.00
Speech0.00

Ownership

FUJITSU LIMITED

assignment · 479240036

OKINAWA INSTITUTE OF SCIENCE AND TECHNOLOGY SCHOOL CORPORATION

assignment · 479240036

Assignors

SASAKI, TOMOTAKE, UCHIBE, EIJI, DOYA, KENJI, ANAI, HIROKAZU, YANAMI, HITOSHI, IWANE, HIDENAO

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC