Patent №
US 10,839,822
Granted
2020-11-17
Filed 2017
Owner
MICROSOFT TECHNOLOGY LICENSING, LLC
Lab
AI components
4
ml · nlp · speech · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
15805106
Representative embodiments disclose mechanisms to separate and recognize multiple audio sources (e.g., picking out individual speakers) in an environment where they overlap and interfere with each other. The architecture uses a microphone array to spatially separate out the audio signals. The spatially filtered signals are then input into a plurality of separators, so each signal is input into a corresponding signal. The separators use neural networks to separate out audio sources. The separators typically produce multiple output signals for the single input signals. A post selection processor then assesses the separator outputs to pick the signals with the highest quality output. These signals can be used in a variety of systems such as speech recognition, meeting transcription and enhancement, hearing aids, music information retrieval, speech enhancement and so forth.
AI classification
Ownership
MICROSOFT TECHNOLOGY LICENSING, LLC
assignment · 440450677
Assignors
CHEN, ZHUO, GONG, YIFAN, WANG, HUAMING, LI, JINYU, XIAO, XIONG, YOSHIOKA, TAKUYA, WANG, ZHENGHAO
On an employer assignment, the assignors are typically the inventors.