Patent №
US 11,699,456
Granted
2023-07-11
Filed 2021
Owner
—
Lab
—
AI components
2
nlp · speech
Assignment
None on record
Dataset
AIPD
2023_r1 edition
Application
17175246
Systems and methods are described for generating a transcript of a legal proceeding or other multi-speaker conversation or performance in real time or near-real time using multi-channel audio capture. Different speakers or participants in a conversation may each be assigned a separate microphone that is placed in proximity to the given speaker, where each audio channel includes audio captured by a different microphone. Filters may be applied to isolate each channel to include speech utterances of a different speaker, and these filtered channels of audio data may then be processed in parallel to generate speech-to-text results that are interleaved to form a generated transcript.