SYSTEM AND METHOD FOR RENDERING THREE DIMENSIONAL FACE MODEL BASED ON AUDIO STREAM AND IMAGE DATA

Patent №

US 11,113,859

Granted

2021-09-07

Filed 2019

Owner

FACEBOOK TECHNOLOGIES, LLC

AI components

3

vision · speech · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16507862

Disclosed herein includes a system, a method, and a non-transitory computer readable medium for rendering a three-dimensional (3D) model of an avatar according to an audio stream including a vocal output of a person and image data capturing a face of the person. In one aspect, phonemes of the vocal output are predicted according to the audio stream, and the predicted phonemes of the vocal output are translated into visemes. In one aspect, a plurality of blendshapes and corresponding weights are determined, according to the corresponding image data of the face, to form the 3D model of the avatar of the person. The visemes may be combined with the 3D model of the avatar to form a 3D representation of the avatar, by synchronizing the visemes with the 3D model of the avatar in time.

VisionSpeechAI hardwareG06T 13/205G06F 40/44G06T 13/40G06T 15/005G06T 17/205G10L 21/10G06T 2210/44G10L 2015/025

AI classification

Vision1.00
Speech1.00
AI hardware0.95
Natural language0.44
Knowledge representation0.16
Machine learning0.11
Evolutionary computation0.08
Planning0.00

Ownership

FACEBOOK TECHNOLOGIES, LLC

assignment · 506430139

Assignors

XIAO, TONG, FU, SIDI, LIU, MENGQIAN, GUO, PEIHONG, LIANG, SHU, ZATEPYAKIN, EVGENY

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC