LEXICOGRAPHIC DEEP REINFORCEMENT LEARNING USING STATE CONSTRAINTS AND CONDITIONAL POLICIES

Patent №

US 11,410,023

Granted

2022-08-09

Filed 2019

Owner

INTERNATIONAL BUSINESS MACHINES CORPORATION

AI components

5

ml · vision · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16290413

A computer-implemented method is provided for modified Lexicographic Reinforcement Learning. The computer implemented method includes obtaining, by a hardware processor, a sequence of tasks. Each of the tasks corresponds to, and has a one-to-one correspondence with, a respective award from among set of rewards. The method further includes performing, by the hardware processor for each of the tasks, reinforcement learning and deep learning for both of (i) one or more policies and (ii) one or more value functions, with a plurality of sets of samples. A plurality of solutions in a form of the one or more policies and the one or more value functions are parametrized by a single neural network with a selector which selects an input of the single neural network from among the plurality of sets of samples.

Machine learningVisionKnowledge representationPlanningAI hardwareG06N 3/048G06N 3/08G06N 3/04G06N 3/0499G06N 3/082G06N 3/092G06N 3/096

AI classification

Machine learning1.00
Planning1.00
AI hardware0.99
Knowledge representation0.86
Vision0.73
Evolutionary computation0.14
Natural language0.04
Speech0.00

Ownership

INTERNATIONAL BUSINESS MACHINES CORPORATION

assignment · 484830238

Assignors

AGRAVANTE, DON JOVEN R., MUNAWAR, ASIM, TACHIBANA, RYUKI

On an employer assignment, the assignors are typically the inventors.

From the same owner

© 2026 NYSGPT2525 LLC