TRAINABLE VIDEOREALISTIC SPEECH ANIMATION

Patent №

US 7,168,953

Granted

2007-01-30

Filed 2003

Owner

MASSACHUSETTS INSTITUTE OF TECHNOLOGY

AI components

6

ml · nlp · vision · speech · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

10352319

A method and apparatus for videorealistic, speech animation is disclosed. A human subject is recorded using a video camera as he/she utters a predetermined speech corpus. After processing the corpus automatically, a visual speech module is learned from the data that is capable of synthesizing the human subject's mouth uttering entirely novel utterances that were not recorded in the original video. The synthesized utterance is re-composited onto a background sequence which contains natural head and eye movement. The final output is videorealistic in the sense that it looks like a video camera recording of the subject. The two key components of this invention are 1) a multidimensional morphable model (MMM) to synthesize new, previously unseen mouth configurations from a small set of mouth image prototypes; and 2) a trajectory synthesis technique based on regularization, which is automatically trained from the recorded video corpus, and which is capable of synthesizing trajectories in MMM space corresponding to any desired utterance.

AI classification

Vision1.00
Speech1.00
Natural language1.00
Machine learning1.00
AI hardware0.91
Planning0.51
Knowledge representation0.00
Evolutionary computation0.00

Ownership

MASSACHUSETTS INSTITUTE OF TECHNOLOGY

assignment · 137140102

Assignors

POGGIO, TOMASO A., EZZAT. ANTOINE F.

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC