FAST DEEP NEURAL NETWORK FEATURE TRANSFORMATION VIA OPTIMIZED MEMORY BANDWIDTH UTILIZATION

Patent №

US 10,013,652

Granted

2018-07-03

Filed 2015

Owner

NUANCE COMMUNICATIONS, INC.

Lab

AI components

3

ml · kr · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

14699778

Deep Neural Networks (DNNs) with many hidden layers and many units per layer are very flexible models with a very large number of parameters. As such, DNNs are challenging to optimize. To achieve real-time computation, embodiments disclosed herein enable fast DNN feature transformation via optimized memory bandwidth utilization. To optimize memory bandwidth utilization, a rate of accessing memory may be reduced based on a batch setting. A memory, corresponding to a selected given output neuron of a current layer of the DNN, may be updated with an incremental output value computed for the selected given output neuron as a function of input values of a selected few non-zero input neurons of a previous layer of the DNN in combination with weights between the selected few non-zero input neurons and the selected given output neuron, wherein a number of the selected few corresponds to the batch setting.

Machine learningKnowledge representationAI hardwareG06N 3/08G06N 3/045G06N 3/0495G06N 3/0499G10L 15/16G10L 15/02G10L 2015/0635

AI classification

Machine learning1.00
AI hardware1.00
Knowledge representation1.00
Vision0.11
Planning0.03
Speech0.00
Evolutionary computation0.00
Natural language0.00

Ownership

NUANCE COMMUNICATIONS, INC.

assignment · 361610422

Assignors

VLIETINCK, JAN, KANTHAK, STEPHAN, VUERINCKX, RUDI, RIS, CHRISTOPHE

On an employer assignment, the assignors are typically the inventors.

From the same owner

© 2026 NYSGPT2525 LLC