SPEECH RECOGNIZER WITH SEGMENT-BASED SIMILARITY FOR LOW COMPLEXITY

Patent №

US 6,230,129

Granted

2001-05-08

Filed 1998

Owner

MATSUSHITA ELECTRIC INDUSTRIAL CO., LTD.

Lab

AI components

3

ml · nlp · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

09199721

A digital word prototype is constructed using one or more speech utterance for a given spoken word or phrase. First, a phone model is used to derive phoneme similarity time series for each of a plurality of phonemes which represent the degree of similarity between the speech utterance and a set of standard phonemes contained in the phone model. Next, the phoneme similarity data is normalized in relation to a non-speech part of the input speech signal. The normalized phoneme similarity data is divided into segments, such that the sum of all normalized phoneme similarity values in a segment are equal for each segment. Next, a word model is constructed from the phoneme similarity data. To do so, within each segment, a summation value is determined by summing over speech frames each of the normalized phoneme similarity values associated with a particular phoneme. In this way, the word model is represented by a vector of summation values that compactly correlate to the normalized phoneme similarity data. Lastly, the results of the individually processed utterances for a given spoken word (i.e., the individual word models) are combined to produce a digital word prototype that electronically represents the given spoken word.

AI classification

Natural language1.00
Speech1.00
Machine learning1.00
AI hardware0.13
Vision0.05
Knowledge representation0.01
Evolutionary computation0.00
Planning0.00

Ownership

MATSUSHITA ELECTRIC INDUSTRIAL CO., LTD.

assignment · 97770989

Assignors

MORIN, PHILIPPE, APPLEBAUM, TED

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC