HYBRID COMPRESSION OF TEXT-TO-SPEECH VOICE DATA

Patent №

US 9,064,489

Granted

2015-06-23

Filed 2012

Owner

IVONA SOFTWARE SP. Z.O.O.

Lab

AI components

2

nlp · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

13720900

Recorded or synthesized speech segments of text-to-speech (TTS) systems may be compressed though the use of both time domain compression and perceptual compression techniques. The twice-compressed recording may be separated into speech segments corresponding to words or subword units for use in a TTS system. The compression rate of time domain compression, and the ratio of time domain compression to perceptual compression, may be modified for any speech segment. The compression amount or ratio may be determined based on linguistic or acoustic features of the word or subword unit that the speech segment represents. Differing compression amounts and ratios may be applied to portions of a single speech segment.

Natural languageSpeechG10L 13/04G10L 19/00G10L 21/04

AI classification

Speech1.00
Natural language1.00
Machine learning0.02
AI hardware0.00
Knowledge representation0.00
Evolutionary computation0.00
Vision0.00
Planning0.00

Ownership

IVONA SOFTWARE SP. Z.O.O.

assignment · 301280226

Assignors

KASZCZUK, MICHAL T., OSOWSKI, LUKASZ M.

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC