BATCH PROCESSING OF REQUESTS FOR TRAINED MACHINE LEARNING MODEL

Patent №

US 10,846,096

Granted

2020-11-24

Filed 2018

Owner

A9.COM, INC.

AI components

3

ml · kr · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16132750

Memory management is provided for processors, such as GPUs used to process data using a trained machine learning model. Requests received to a CPU can be stored to a request queue until the queue is full, or until a timeout value has been reached for periods of lower activity. The requests can then be batched and sent to a GPU as a single message on a single thread. Memory can be pre-allocated, and the trained model loaded into GPU memory once for processing of the relevant batches. The individual requests can be processed by the GPU and the results analyzed to determine at least a subset of results to return to the CPU, which can be provided back as results of the processing.

Machine learningKnowledge representationAI hardwareG06F 9/3856G06F 9/5083G06F 3/0659G06F 9/5016G06N 20/00G06T 1/20G06T 1/60

AI classification

AI hardware1.00
Machine learning0.98
Knowledge representation0.59
Vision0.02
Evolutionary computation0.01
Natural language0.01
Speech0.00
Planning0.00

Ownership

A9.COM, INC.

assignment · 468880279

Assignors

CHUNG, KIUK, KANDROT, EDWARD, LE GRAND, SCOTT MICHAEL

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC