DEEP LANGUAGE AND ACOUSTIC MODELING CONVERGENCE AND CROSS TRAINING

Patent №

US 11,270,686

Granted

2022-03-08

Filed 2017

Owner

INTERNATIONAL BUSINESS MACHINES CORPORATION

AI components

3

ml · nlp · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

15471436

A model-pair is selected to recognize spoken words in a speech signal generated from a speech, which includes an acoustic model and a language model. A degree of disjointedness between the acoustic model and the language model is computed relative to the speech by comparing a first recognition output produced from the acoustic model and a second recognition output produced from the language model. When the acoustic model incorrectly recognizes a portion of the speech signal as a first word and the language model correctly recognizes the portion of the speech signal as a second word, a textual representation of the second word is determined and associated with a set of sound descriptors to generate a training speech pattern. Using the training speech pattern, the acoustic model is trained to recognize the portion of the speech signal as the second word.

Machine learningNatural languageSpeechG10L 15/063G10L 15/01G10L 15/16G10L 15/1807G10L 15/183G10L 15/30G10L 25/51

AI classification

Natural language1.00
Speech1.00
Machine learning1.00
AI hardware0.29
Knowledge representation0.23
Evolutionary computation0.00
Vision0.00
Planning0.00

Ownership

INTERNATIONAL BUSINESS MACHINES CORPORATION

assignment · 417660891

Assignors

BAUGHMAN, AARON K., GANCI, JOHN M., JR, HAMMER, STEPHEN C., TRIM, CRAIG M.

On an employer assignment, the assignors are typically the inventors.

From the same owner

© 2026 NYSGPT2525 LLC