System and Method for Speech Recognition Using Pitch-Synchronous Spectral Parameters
Patent №
US 8,942,977
Granted
2015-01-27
Filed 2014
Owner
—
Lab
—
AI components
3
ml · nlp · speech
Assignment
None on record
Dataset
AIPD
2023_r1 edition
Application
14216684
The present invention defines a pitch-synchronous parametrical representation of speech signals as the basis of speech recognition, and discloses methods of generating the said pitch-synchronous parametrical representation from speech signals. The speech signal is first going through a pitch-marks picking program to identify the pitch periods. The speech signal is then segmented into pitch-synchronous frames. An ends-matching program equalizes the values at the two ends of the waveform in each frame. Using Fourier analysis, the speech signal in each frame is converted into a pitch-synchronous amplitude spectrum. Using Laguerre functions, the said amplitude spectrum is converted into a unit vector, referred to as the timbre vector. By using a database of correlated phonemes and timbre vectors, the most likely phoneme sequence of an input speech signal can be decoded in the acoustic stage of a speech recognition system.