Discriminative Training of Document Transcription System

Patent №

US 9,520,124

Granted

2016-12-13

Filed 2015

Owner

MMODAL IP LLC

Lab

AI components

6

ml · nlp · speech · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

14942349

A system is provided for training an acoustic model for use in speech recognition. In particular, such a system may be used to perform training based on a spoken audio stream and a non-literal transcript of the spoken audio stream. Such a system may identify text in the non-literal transcript which represents concepts having multiple spoken forms. The system may attempt to identify the actual spoken form in the audio stream which produced the corresponding text in the non-literal transcript, and thereby produce a revised transcript which more accurately represents the spoken audio stream. The revised, and more accurate, transcript may be used to train the acoustic model using discriminative training techniques, thereby producing a better acoustic model than that which would be produced using conventional techniques, which perform training based directly on the original non-literal transcript.

Machine learningNatural languageSpeechKnowledge representationPlanningAI hardwareG10L 15/063G06F 40/211G06F 40/289G06F 40/40G10L 15/02G10L 15/183G10L 15/26G16H 15/00+3 more

AI classification

Natural language1.00
Speech1.00
Machine learning1.00
Planning0.99
AI hardware0.99
Knowledge representation0.72
Vision0.00
Evolutionary computation0.00

Ownership

MMODAL IP LLC

assignment · 386070177

Assignors

MATHIAS, LAMBERT, YEGNANARAYANAN, GIRIJA, FRITSCH, JUERGEN

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC