DUAL MODEL SPEAKER IDENTIFICATION

Patent №

US 9,711,148

Granted

2017-07-18

Filed 2013

Owner

GOOGLE INC.

AI components

5

ml · nlp · vision · speech · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

13944975

A processing system receives an audio signal encoding an utterance and determines that a first portion of the audio signal corresponds to a predefined phrase. The processing system accesses one or more text-dependent models associated with the predefined phrase and determines a first confidence based on the one or more text-dependent models associated with the predefined phrase, the first confidence corresponding to a first likelihood that a particular speaker spoke the utterance. The processing system determines a second confidence for a second portion of the audio signal using one or more text-independent models, the second confidence corresponding to a second likelihood that the particular speaker spoke the utterance. The processing system then determines that the particular speaker spoke the utterance based at least in part on the first confidence and the second confidence.

Machine learningNatural languageVisionSpeechAI hardwareG10L 17/24G10L 15/02G10L 15/22G10L 17/02G10L 17/04G10L 17/10G10L 17/22

AI classification

Natural language1.00
Speech1.00
Machine learning1.00
Vision0.96
AI hardware0.86
Knowledge representation0.40
Evolutionary computation0.03
Planning0.00

Ownership

GOOGLE INC.

assignment · 313400095

Assignors

SHARIFI, MATTHEW, ROBLEK, DOMINIK

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC