Altering Audio to Improve Automatic Speech Recognition

Patent №

US 11,488,591

Granted

2022-11-01

Filed 2019

Owner

AMAZON TECHNOLOGIES, INC.

+1 more

AI components

3

nlp · speech · planning

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16510060

Techniques for altering audio being output by a voice-controlled device, or another device, to enable more accurate automatic speech recognition (ASR) by the voice-controlled device. For instance, a voice-controlled device may output audio within an environment using a speaker of the device. While outputting the audio, a microphone of the device may capture sound within the environment and may generate an audio signal based on the captured sound. The device may then analyze the audio signal to identify speech of a user within the signal, with the speech indicating that the user is going to provide a subsequent command to the device. Thereafter, the device may alter the output of the audio (e.g., attenuate the audio, pause the audio, switch from stereo to mono, etc.) to facilitate speech recognition of the user's subsequent command.

Natural languageSpeechPlanningG10L 15/22G10L 15/20G10L 17/00G11B 27/005H03G 3/32H03G 5/02H04R 3/12G10L 15/26+1 more

AI classification

Speech1.00
Natural language1.00
Planning0.97
AI hardware0.17
Vision0.13
Knowledge representation0.03
Machine learning0.01
Evolutionary computation0.01

Ownership

AMAZON TECHNOLOGIES, INC.

assignment · 497370471

RAWLES LLC

assignment · 497370781

Assignors

HART, GREGORY M., WORLEY, WILLIAM SPENCER, III

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC