Patent №
US 5,774,837
Granted
1998-06-30
Filed 1995
Owner
VOXWARE, INC.
Lab
—
AI components
1
speech
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
08528513
A modular system and method is provided for encoding and decoding of speech signals using voicing probability determination. The continuous input speech is divided into time segments of a predetermined length. For each segment the encoder of the system computes the signal pitch and a parameter which is related to the relative content of voiced and unvoiced portions in the spectrum of the signal, which is expressed as a ratio Pv, defined as a voicing probability. The voiced portion of the signal spectrum, as determined by the parameter Pv, is encoded using a set of harmonically related amplitudes corresponding to the estimated pitch. The unvoiced portion of the signal is processed in a separate processing branch which uses a modified linear predictive coding algorithm. Parameters representing both the voiced and the unvoiced portions of a speech segment are combined in data packets for transmission. In the decoder, speech is synthesized from the transmitted parameters representing voiced and unvoiced portions of the speech in a reverse order. Boundary conditions between voiced and unvoiced segments are established to ensure amplitude and phase continuity for improved output speech quality. Perceptually smooth transition between frames is ensured by using an overlap and add method of synthesis. Also disclosed is the use of the system in the generation of a variety of voice effects.
AI classification
Ownership
VOXWARE, INC.
assignment · 77520910
Assignors
YELDENER, SUAT, AGUILAR, JOSEPH GERARD
On an employer assignment, the assignors are typically the inventors.