SPEECH SIGNAL SEPARATION AND SYNTHESIS BASED ON AUDITORY SCENE ANALYSIS AND SPEECH MODELING

Patent №

US 9,536,540

Granted

2017-01-03

Filed 2014

Owner

AUDIENCE, INC.

Lab

AI components

3

ml · nlp · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

14335850

Provided are systems and methods for generating clean speech from a speech signal representing a mixture of a noise and speech. The clean speech may be generated from synthetic speech parameters. The synthetic speech parameters are derived based on the speech signal components and a model of speech using auditory and speech production principles. The modeling may utilize a source-filter structure of the speech signal. One or more spectral analyzes on the speech signal are performed to generate spectral representations. The feature data is derived based on a spectral representation. The features corresponding to the target speech according to a model of speech are grouped and separated from the feature data. The synthetic speech parameters, including spectral envelope, pitch data and voice classification data are generated based on features corresponding to the target speech.

AI classification

Speech1.00
Natural language1.00
Machine learning1.00
AI hardware0.32
Vision0.08
Evolutionary computation0.00
Knowledge representation0.00
Planning0.00

Ownership

AUDIENCE, INC.

assignment · 357150433

© 2026 NYSGPT2525 LLC