Patent №
US 8,335,688
Granted
2012-12-18
Filed 2004
Owner
MULTIMODAL TECHNOLOGIES, INC.
+1 more
Lab
—
AI components
5
ml · nlp · speech · planning · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
10922513
A system is provided for training an acoustic model for use in speech recognition. In particular, such a system may be used to perform training based on a spoken audio stream and a non-literal transcript of the spoken audio stream. Such a system may identify text in the non-literal transcript which represents concepts having multiple spoken forms. The system may attempt to identify the actual spoken form in the audio stream which produced the corresponding text in the non-literal transcript, and thereby produce a revised transcript which more accurately represents the spoken audio stream. The revised, and more accurate, transcript may be used to train the acoustic model, thereby producing a better acoustic model than that which would be produced using conventional techniques, which perform training based directly on the original non-literal transcript.
AI classification
Ownership
MULTIMODAL TECHNOLOGIES, INC.
assignment · 154140831
MMODAL IP LLC
assignment · 314380241
Assignors
YEGNANARAYANAN, GIRIJA, FINKE, MICHAEL, FRITSCH, JUERGEN, KOLL, DETLEF, WOSZCZYNA, MONIKA
On an employer assignment, the assignors are typically the inventors.