TEXT-TO-SPEECH PROCESSING USING PREVIOUSLY SPEECH PROCESSED DATA

Patent №

US 10,140,973

Granted

2018-11-27

Filed 2016

Owner

AMAZON TECHNOLOGIES, INC.

AI components

6

ml · nlp · speech · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

15266116

Systems, methods, and devices for generating text-to-speech output using previously captured speech are described. Spoken audio is obtained and undergoes speech processing to create text. The resulting text is stored with the spoken audio, with both the text and the spoken audio being associated with the individual that spoke the audio. Various spoken audio and corresponding text are stored over time to create a library of speech units. When the individual sends a text message to a recipient, the text message is processed to determine portions of text, and the portions of text are compared to the library of text associated with the individual. When text in the library is identified, the system selects the spoken audio units associated with the identified stored text. The selected spoken audio units are then used to generate output audio data corresponding to the original text message, with the output audio data being sent to a device of the message recipient.

Machine learningNatural languageSpeechKnowledge representationPlanningAI hardwareG10L 13/10G06F 40/247G06F 40/30G06N 7/01G10L 13/07G10L 15/26G10L 13/06

AI classification

Natural language1.00
Speech1.00
Planning0.99
AI hardware0.82
Knowledge representation0.71
Machine learning0.61
Evolutionary computation0.03
Vision0.03

Ownership

AMAZON TECHNOLOGIES, INC.

assignment · 438820084

Assignors

DALMIA, MANISH KUMAR, KUKLINSKI, RAFAL

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC