INSERTING BREATH SOUNDS INTO TEXT-TO-SPEECH OUTPUT

Patent №

US 9,508,338

Granted

2016-11-29

Filed 2013

Owner

AMAZON TECHNOLOGIES, INC.

AI components

2

nlp · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

14081233

A text-to-speech (TTS) system may be configured to incorporate breath sounds in the output speech. By incorporating breath sounds into speech output from text a TTS system may be able to mimic more naturally sounding human speech, particularly for long-form narration of text longer than short phrases. The breath sounds may be stored as units for unit selection or may be generated during parametric synthesis. The acoustic features of the breath sounds and duration between breaths may depend upon the punctuation of text, the linguistic distance between breaths, the breaks between intonational phrases, the linguistic context of the breaths, and other factors.

Natural languageSpeechG10L 13/02G10L 13/06G10L 2013/083

AI classification

Natural language1.00
Speech1.00
Knowledge representation0.01
AI hardware0.01
Evolutionary computation0.00
Machine learning0.00
Planning0.00
Vision0.00

Ownership

AMAZON TECHNOLOGIES, INC.

assignment · 333200541

Assignors

KASZCZUK, MICHAL TADEUSZ, TEGI, MACIEJ, CZUCZMAN, MICHAL, MOIS, REMUS RAZVAN

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC