METHOD AND SYSTEM FOR ENHANCING A SPEECH SIGNAL OF A HUMAN SPEAKER IN A VIDEO USING VISUAL INFORMATION
Patent №
US 10,475,465
Granted
2019-11-12
Filed 2018
Owner
YISSUM RESEARCH DEVELOPMENT COMPANY OF THE HEBREW UNIVERSITY OF JERUSALEM LTD.
Lab
—
AI components
4
ml · nlp · vision · speech
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
16026449
A method and system for enhancing a speech signal is provided herein. The method may include the following steps: obtaining an original video, wherein the original video includes a sequence of original input images showing a face of at least one human speaker, and an original soundtrack synchronized with said sequence of images; and processing, using a computer processor, the original video, to yield an enhanced speech signal of said at least one human speaker, by detecting sounds that are acoustically unrelated to the speech of the at least one human speaker, based on visual data derived from the sequence of original input images.
AI classification
Ownership
YISSUM RESEARCH DEVELOPMENT COMPANY OF THE HEBREW UNIVERSITY OF JERUSALEM LTD.
assignment · 470870262
Assignors
PELEG, SHMUEL, SHAMIR, ASAPH, HALPERIN, TAVI, GABBAY, AVIV, EPHRAT, ARIEL
On an employer assignment, the assignors are typically the inventors.