SPEECH SIGNAL SEPARATION AND SYNTHESIS BASED ON AUDITORY SCENE ANALYSIS AND SPEECH MODELING
Patent №
US 9,536,540
Granted
2017-01-03
Filed 2014
Owner
AUDIENCE, INC.
Lab
—
AI components
3
ml · nlp · speech
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
14335850
Provided are systems and methods for generating clean speech from a speech signal representing a mixture of a noise and speech. The clean speech may be generated from synthetic speech parameters. The synthetic speech parameters are derived based on the speech signal components and a model of speech using auditory and speech production principles. The modeling may utilize a source-filter structure of the speech signal. One or more spectral analyzes on the speech signal are performed to generate spectral representations. The feature data is derived based on a spectral representation. The features corresponding to the target speech according to a model of speech are grouped and separated from the feature data. The synthetic speech parameters, including spectral envelope, pitch data and voice classification data are generated based on features corresponding to the target speech.
AI classification
Ownership
AUDIENCE, INC.
assignment · 357150433