SYSTEM AND METHOD FOR DISCRIMINATIVE PRONUNCIATION MODELING FOR VOICE SEARCH

Patent №

US 8,296,141

Granted

2012-10-23

Filed 2008

Owner

AT&T INTELLECTUAL PROPERTY I, L.P.

Lab

AI components

5

ml · nlp · vision · speech · kr

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

12274025

Disclosed herein are systems, computer-implemented methods, and computer-readable media for speech recognition. The method includes receiving speech utterances, assigning a pronunciation weight to each unit of speech in the speech utterances, each respective pronunciation weight being normalized at a unit of speech level to sum to 1, for each received speech utterance, optimizing the pronunciation weight by (1) identifying word and phone alignments and corresponding likelihood scores, and (2) discriminatively adapting the pronunciation weight to minimize classification errors, and recognizing additional received speech utterances using the optimized pronunciation weights. A unit of speech can be a sentence, a word, a context-dependent phone, a context-independent phone, or a syllable. The method can further include discriminatively adapting pronunciation weights based on an objective function. The objective function can be maximum mutual information (MMI), maximum likelihood (MLE) training, minimum classification error (MCE) training, or other functions known to those of skill in the art. Speech utterances can be names. The speech utterances can be received as part of a multimodal search or input. The step of discriminatively adapting pronunciation weights can further include stochastically modeling pronunciations.

AI classification

Natural language1.00
Speech1.00
Machine learning1.00
Vision0.80
Knowledge representation0.58
AI hardware0.02
Planning0.00
Evolutionary computation0.00

Ownership

AT&T INTELLECTUAL PROPERTY I, L.P.

assignment · 218600959

Assignors

GILBERT, MAZIN, CONKIE, ALISTAIR D., LJOLJE, ANDREJ

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC