METHOD AND APPARATUS FOR LANGUAGE AND SPEAKER RECOGNITION

Patent №

US 5,189,727

Granted

1993-02-23

Filed 1991

Owner

WARFAR ASSOCIATES, INC.

Lab

AI components

3

ml · nlp · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

07752898

An initial learning phase creates histograms for each of the languages to be recognized. A first pass enters a number of samples of speech, and at each predetermined instant of time, each sample of speech is Fast Fourier Transformed (FFT) to create a spectrum showing frequency content of the speech at that instant of time (a spectral vector). The frequency content is compared with frequency contents which have been previously stored. If the current spectral vector is close enough to a previously stored spectral vector, a weighted average between the two is formed, and a weight indicating frequency of occurrence is incremented. If the current value is not similar to one which has been previously stored, it is stored with an initial weight of "1". The most common frequency spectra are determined for all of the languages grouped together to form a composite basis set. A second pass then puts a sample of sounds through the Fast Fourier Transform to again obtain frequency spectrums. The obtained frequency spectrums are compared against all of the prestored frequency spectra in the composite basis set, and a closest match is determined. A number of occurrences of each frequency spectra in the composite basis set is plotted as a histogram. This histogram is used during the recognition phase to determine a closest fit between an unknown language and one of the known languages.

AI classification

Natural language1.00
Speech1.00
Machine learning0.99
AI hardware0.15
Vision0.00
Knowledge representation0.00
Evolutionary computation0.00
Planning0.00

Ownership

WARFAR ASSOCIATES, INC.

assignment · 62590575

Assignors

GUERRERI, STEPHEN J.

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC