TRAINING AND/OR USING A LANGUAGE SELECTION MODEL FOR AUTOMATICALLY DETERMINING LANGUAGE FOR SPEECH RECOGNITION OF SPOKEN UTTERANCE
Patent №
US 11,646,011
Granted
2023-05-09
Filed 2022
Owner
GOOGLE LLC
AI components
4
ml · nlp · speech · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
17846287
Methods and systems for training and/or using a language selection model for use in determining a particular language of a spoken utterance captured in audio data. Features of the audio data can be processed using the trained language selection model to generate a predicted probability for each of N different languages, and a particular language selected based on the generated probabilities. Speech recognition results for the particular language can be utilized responsive to selecting the particular language of the spoken utterance. Many implementations are directed to training the language selection model utilizing tuple losses in lieu of traditional cross-entropy losses. Training the language selection model utilizing the tuple losses can result in more efficient training and/or can result in a more accurate and/or robust model—thereby mitigating erroneous language selections for spoken utterances.
AI classification
Ownership
GOOGLE LLC
assignment · 602770292
Assignors
WAN, LI, YU, YANG, SRIDHAR, PRASHANT, LOPEZ MORENO, IGNACIO, WANG, QUAN
On an employer assignment, the assignors are typically the inventors.