Using Context Information With End-to-End Models for Speech Recognition

Patent №

US 11,545,142

Granted

2023-01-03

Filed 2020

Owner

GOOGLE LLC

AI components

4

ml · nlp · speech · kr

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

16827937

A method includes receiving audio data encoding an utterance, processing, using a speech recognition model, the audio data to generate speech recognition scores for speech elements, and determining context scores for the speech elements based on context data indicating a context for the utterance. The method also includes executing, using the speech recognition scores and the context scores, a beam search decoding process to determine one or more candidate transcriptions for the utterance. The method also includes selecting a transcription for the utterance from the one or more candidate transcriptions.

Machine learningNatural languageSpeechKnowledge representationG10L 15/16G06F 18/2113G06N 3/0442G06N 3/0455G06N 3/08G06N 3/0895G06N 3/09G06N 20/00+6 more

AI classification

Natural language1.00
Speech1.00
Machine learning1.00
Knowledge representation0.99
AI hardware0.00
Vision0.00
Evolutionary computation0.00
Planning0.00

Ownership

GOOGLE LLC

assignment · 523390553

Assignors

ZHAO, DING, LI, BO, PANG, RUOMING, SAINATH, TARA N., RYBACH, DAVID, BHATIA, DEEPTI, WU, ZELIN

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC