COARTICULATION METHOD FOR AUDIO-VISUAL TEXT-TO-SPEECH SYNTHESIS

Patent №

US 6,112,177

Granted

2000-08-29

Filed 1997

Owner

AT&T CORP.

Lab

AI components

4

ml · nlp · vision · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

08965702

A method for generating animated sequences of talking heads in text-to-speech applications wherein a processor samples a plurality of frames comprising image samples. Representative parameters are extracted from the image samples and stored in an animation library. The processor also samples a plurality of multiphones comprising images together with their associated sounds. The processor extracts parameters from these images comprising data characterizing mouth shapes, maps, rules, or equations, and stores the resulting parameters and sound information in a coarticulation library. The animated sequence begins with the processor considering an input phoneme sequence, recalling from the coarticulation library parameters associated with that sequence, and selecting appropriate image samples from the animation library based on that sequence. The image samples are concatenated together, and the corresponding sound is output, to form the animated synthesis.

AI classification

Speech1.00
Natural language1.00
Machine learning0.98
Vision0.95
AI hardware0.21
Knowledge representation0.00
Evolutionary computation0.00
Planning0.00

Ownership

AT&T CORP.

assignment · 93770802

Assignors

COSATTO, ERIC, GRAF, HANS PETER, POTAMIANOS, GERASIMOS

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC