SPEECH AND TEXT DRIVEN HMM-BASED BODY ANIMATION SYNTHESIS

Patent №

US 8,224,652

Granted

2012-07-17

Filed 2008

Owner

MICROSOFT CORPORATION

AI components

7

ml · nlp · vision · speech · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

12239564

An “Animation Synthesizer” uses trainable probabilistic models, such as Hidden Markov Models (HMM), Artificial Neural Networks (ANN), etc., to provide speech and text driven body animation synthesis. Probabilistic models are trained using synchronized motion and speech inputs (e.g., live or recorded audio/video feeds) at various speech levels, such as sentences, phrases, words, phonemes, sub-phonemes, etc., depending upon the available data, and the motion type or body part being modeled. The Animation Synthesizer then uses the trainable probabilistic model for selecting animation trajectories for one or more different body parts (e.g., face, head, hands, arms, etc.) based on an arbitrary text and/or speech input. These animation trajectories are then used to synthesize a sequence of animations for digital avatars, cartoon characters, computer generated anthropomorphic persons or creatures, actual motions for physical robots, etc., that are synchronized with a speech output corresponding to the text and/or speech input.

AI classification

Natural language1.00
Speech1.00
Machine learning1.00
AI hardware1.00
Planning0.98
Knowledge representation0.98
Vision0.91
Evolutionary computation0.01

Ownership

MICROSOFT CORPORATION

assignment · 225120164

Assignors

WANG, LIJUAN, MA, LEI, SOONG, FRANK KAO-PING

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC