SYSTEM AND METHOD FOR PROVIDING HIGH-QUALITY STRETCHING AND COMPRESSION OF A DIGITAL AUDIO SIGNAL
Patent №
US 7,337,108
Granted
2008-02-26
Filed 2003
Owner
MICROSOFT CORPORATION
Lab
AI components
1
speech
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
10660325
An adaptive “temporal audio scaler” is provided for automatically stretching and compressing frames of audio signals received across a packet-based network. Prior to stretching or compressing segments of a current frame, the temporal audio scaler first computes a pitch period for each frame for sizing signal templates used for matching operations in stretching and compressing segments. Further, the temporal audio scaler also determines the type or types of segments comprising each frame. These segment types include “voiced” segments, “unvoiced” segments, and “mixed” segments which include both voiced and unvoiced portions. The stretching or compression methods applied to segments of each frame are then dependent upon the type of segments comprising each frame. Further, the amount of stretching and compression applied to particular segments is automatically variable for minimizing signal artifacts while still ensuring that an overall target stretching or compression ratio is maintained for each frame.
AI classification
Ownership
MICROSOFT CORPORATION
assignment · 144960905
Assignors
FLORENCIO, DINEI A., CHOU, PHILIP A., HE, LI-WEI
On an employer assignment, the assignors are typically the inventors.