System and Method for Speech Recognition Using Pitch-Synchronous Spectral Parameters

Patent №

US 8,942,977

Granted

2015-01-27

Filed 2014

Owner

Lab

AI components

3

ml · nlp · speech

Assignment

None on record

Dataset

AIPD

2023_r1 edition

Application

14216684

The present invention defines a pitch-synchronous parametrical representation of speech signals as the basis of speech recognition, and discloses methods of generating the said pitch-synchronous parametrical representation from speech signals. The speech signal is first going through a pitch-marks picking program to identify the pitch periods. The speech signal is then segmented into pitch-synchronous frames. An ends-matching program equalizes the values at the two ends of the waveform in each frame. Using Fourier analysis, the speech signal in each frame is converted into a pitch-synchronous amplitude spectrum. Using Laguerre functions, the said amplitude spectrum is converted into a unit vector, referred to as the timbre vector. By using a database of correlated phonemes and timbre vectors, the most likely phoneme sequence of an input speech signal can be decoded in the acoustic stage of a speech recognition system.

AI classification

Natural language1.00
Speech1.00
Machine learning0.99
Vision0.36
Planning0.05
AI hardware0.00
Evolutionary computation0.00
Knowledge representation0.00
© 2026 NYSGPT2525 LLC