METHOD AND SYSTEM FOR IMPROVING A POLICY FOR A STOCHASTIC CONTROL PROBLEM

Patent №

US 11,017,289

Granted

2021-05-25

Filed 2017

Owner

1QB INFORMATION TECHNOLOGIES INC.

Lab

AI components

6

ml · vision · kr · planning · evo · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

15590614

A method and system for improving a stochastic control problem policy, the method including a sampling device obtaining data representing sample Boltzmann machine configurations, obtaining a stochastic control problem's initialization data and initial policy; assigning representative data of initial coupler weights and node biases and the Boltzmann machine's transverse field strength to the sampling device; until a stopping criterion is met, generating a present-epoch state-action pair, amending, sampling for the present-epoch state-action pair, approximating a present-epoch state-action Q-function value, obtaining a future-epoch state-action pair through a stochastic state process including a stochastic optimization test on all state-action pairs to provide the action at the future-epoch and update the future-epoch state's policy; amending the representative data, sampling for the future-epoch state-action pair, obtaining a future-epoch state-action Q-function value, updating each weight and bias and providing the policy when the stopping criterion is met.

Machine learningVisionKnowledge representationPlanningEvolutionary computationAI hardwareG06N 3/044G06N 3/08G06E 3/005G06N 3/047G06N 3/0475G06N 3/088G06N 3/092G06N 7/01+3 more

AI classification

Planning1.00
Machine learning1.00
AI hardware1.00
Knowledge representation0.99
Evolutionary computation0.85
Vision0.83
Natural language0.00
Speech0.00

Ownership

1QB INFORMATION TECHNOLOGIES INC.

assignment · 425780704

Assignors

CRAWFORD, DANIEL, RONAGH, POOYA, LEVIT, ANNA

On an employer assignment, the assignors are typically the inventors.

From the same owner

© 2026 NYSGPT2525 LLC