METHOD AND APPARATUS FOR PRODUCING AUDIO-VISUAL SYNTHETIC SPEECH

Patent №

US 5,657,426

Granted

1997-08-12

Filed 1994

Owner

DIGITAL EQUIPMENT CORPORATION

Lab

AI components

2

nlp · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

08258145

A method and apparatus provide a video image of facial features synchronized with synthetic speech. Text input is transformed into a string of phonemes and timing data, which are transmitted to an image generation unit. At the same time, a string of synthetic speech samples is transmitted to an audio server. The audio server produces signals for an audio speaker, causing the audio signals to be continuously audibilized; additionally, the audio server initializes a timer. The image generation unit reads the timing data from the timer and, by consulting the phoneme and timing data, determines the position of the phoneme currently being audibilized. The image generation unit then calculates the facial configuration corresponding to the position in the string of phonemes, calculates the facial configuration, and causes the facial configuration to be displayed on a video device.

AI classification

Natural language1.00
Speech1.00
Vision0.18
AI hardware0.05
Machine learning0.01
Evolutionary computation0.01
Planning0.01
Knowledge representation0.00

Ownership

DIGITAL EQUIPMENT CORPORATION

assignment · 70420688

Assignors

WATERS, KEITH, LEVERGOOD, THOMAS M.

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC