AUTOMATED TRANSCRIPT GENERATION FROM MULTI-CHANNEL AUDIO

Patent №

US 11,699,456

Granted

2023-07-11

Filed 2021

Owner

Lab

AI components

2

nlp · speech

Assignment

None on record

Dataset

AIPD

2023_r1 edition

Application

17175246

Systems and methods are described for generating a transcript of a legal proceeding or other multi-speaker conversation or performance in real time or near-real time using multi-channel audio capture. Different speakers or participants in a conversation may each be assigned a separate microphone that is placed in proximity to the given speaker, where each audio channel includes audio captured by a different microphone. Filters may be applied to isolate each channel to include speech utterances of a different speaker, and these filtered channels of audio data may then be processed in parallel to generate speech-to-text results that are interleaved to form a generated transcript.

Natural languageSpeechG10L 21/10G06F 3/165G06F 3/167G10L 17/00G10L 21/0232H04R 1/406H04R 3/005G10L 2021/02082+3 more

AI classification

Speech1.00
Natural language1.00
Planning0.07
Knowledge representation0.04
Machine learning0.03
AI hardware0.02
Vision0.00
Evolutionary computation0.00
© 2026 NYSGPT2525 LLC