FACTORIAL HIDDEN MARKOV MODEL FOR AUDIOVISUAL SPEECH RECOGNITION

Patent №

US 7,209,883

Granted

2007-04-24

Filed 2002

Owner

INTEL CORPORATION

Lab

AI components

4

ml · nlp · vision · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

10142447

A speech recognition method includes use of synchronous or asynchronous audio and a video data to enhance speech recognition probabilities. A two stream factorial hidden Markov model is trained and used to identify speech. At least one stream is derived from audio data and a second stream is derived from mouth pattern data. Gestural or other suitable data streams can optionally be combined to reduce speech recognition error rates in noisy environments.

Machine learningNatural languageVisionSpeechG10L 15/24G06F 18/295G06V 40/20G10L 15/142

AI classification

Natural language1.00
Vision1.00
Speech1.00
Machine learning1.00
AI hardware0.10
Knowledge representation0.03
Planning0.00
Evolutionary computation0.00

Ownership

INTEL CORPORATION

assignment · 132440946

Assignors

NEFIAN, ARA V.

On an employer assignment, the assignors are typically the inventors.

From the same owner

© 2026 NYSGPT2525 LLC