RECOGNITION OF SPEECH IN EDITABLE AUDIO STREAMS

Patent №

US 7,869,996

Granted

2011-01-11

Filed 2007

Owner

MULTIMODAL TECHNOLOGIES, INC.

Lab

AI components

2

nlp · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

11944517

A speech processing system divides a spoken audio stream into partial audio streams, referred to as “snippets.” The system may divide a portion of the audio stream into two snippets at a position at which the speaker performed an editing operation, such as pausing and then resuming recording, or rewinding and then resuming recording. The snippets may be transmitted sequentially to a consumer, such as an automatic speech recognizer or a playback device, as the snippets are generated. The consumer may process (e.g., recognize or play back) the snippets as they are received. The consumer may modify its output in response to editing operations reflected in the snippets. The consumer may process the audio stream while it is being created and transmitted even if the audio stream includes editing operations that invalidate previously-transmitted partial audio streams, thereby enabling shorter turnaround time between dictation and consumption of the complete audio stream.

Natural languageSpeechG10L 15/22G11B 27/036G11B 27/105

AI classification

Speech1.00
Natural language1.00
Knowledge representation0.37
Planning0.33
Vision0.04
Evolutionary computation0.03
AI hardware0.01
Machine learning0.00

Ownership

MULTIMODAL TECHNOLOGIES, INC.

assignment · 202390769

Assignors

CARRAUX, ERIC, KOLL, DETLEF

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC