Patent №
US 10,223,333
Granted
2019-03-05
Filed 2015
Owner
NVIDIA CORPORATION
Lab
AI components
3
ml · vision · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
14838291
In one embodiment of the present invention a convolution engine configures a parallel processing pipeline to perform multi-convolution operations. More specifically, the convolution engine configures the parallel processing pipeline to independently generate and process individual image tiles. In operation, for each image tile, the pipeline calculates source locations included in an input image batch. Notably, the source locations reflect the contribution of the image tile to an output tile of an output matrix—the result of the multi-convolution operation. Subsequently, the pipeline copies data from the source locations to the image tile. Similarly, the pipeline copies data from a filter stack to a filter tile. The pipeline then performs matrix multiplication operations between the image tile and the filter tile to generate data included in the corresponding output tile. To optimize both on-chip memory usage and execution time, the pipeline creates each image tile in on-chip memory as-needed.
AI classification
Ownership
NVIDIA CORPORATION
assignment · 390900361
Assignors
CHETLUR, SHARANYAN, CATANZARO, BRYAN
On an employer assignment, the assignors are typically the inventors.