TEXT AND AUDIO-BASED REAL-TIME FACE REENACTMENT

Patent №

US 11,114,086

Granted

2021-09-07

Filed 2019

Owner

AI FACTORY, INC.

Lab

AI components

5

ml · nlp · vision · speech · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16509370

Provided are systems and methods for text and audio-based real-time face reenactment. An example method includes receiving an input text and a target image, the target image including a target face; generating, based on the input text, a sequence of sets of acoustic features representing the input text; determining, based on the sequence of sets of acoustic features, a sequence of sets of scenario data indicating modifications of the target face for pronouncing the input text; generating, based on the sequence of sets of scenario data, a sequence of frames, wherein each of the frames includes the target face modified based on at least one of the sets of scenario data; generating, based on the sequence of frames, an output video; and synthesizing, based on the sequence of sets of acoustic features, an audio data and adding the audio data to the output video.

Machine learningNatural languageVisionSpeechAI hardwareG10L 13/00G06T 13/40G06V 10/764G06V 10/82G06V 40/171G10L 13/08

AI classification

Natural language1.00
Vision1.00
Speech1.00
AI hardware0.99
Machine learning0.61
Planning0.36
Knowledge representation0.00
Evolutionary computation0.00

Ownership

AI FACTORY, INC.

assignment · 497370359

Assignors

SAVCHENKOV, PAVEL, LUKIN, MAXIM, MASHRABOV, ALEKSANDR

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC