SYSTEM AND METHOD FOR RENDERING THREE DIMENSIONAL FACE MODEL BASED ON AUDIO STREAM AND IMAGE DATA
Patent №
US 11,113,859
Granted
2021-09-07
Filed 2019
Owner
FACEBOOK TECHNOLOGIES, LLC
AI components
3
vision · speech · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
16507862
Disclosed herein includes a system, a method, and a non-transitory computer readable medium for rendering a three-dimensional (3D) model of an avatar according to an audio stream including a vocal output of a person and image data capturing a face of the person. In one aspect, phonemes of the vocal output are predicted according to the audio stream, and the predicted phonemes of the vocal output are translated into visemes. In one aspect, a plurality of blendshapes and corresponding weights are determined, according to the corresponding image data of the face, to form the 3D model of the avatar of the person. The visemes may be combined with the 3D model of the avatar to form a 3D representation of the avatar, by synchronizing the visemes with the 3D model of the avatar in time.
AI classification
Ownership
FACEBOOK TECHNOLOGIES, LLC
assignment · 506430139
Assignors
XIAO, TONG, FU, SIDI, LIU, MENGQIAN, GUO, PEIHONG, LIANG, SHU, ZATEPYAKIN, EVGENY
On an employer assignment, the assignors are typically the inventors.