VOICE TAGGING, VOICE ANNOTATION, AND SPEECH RECOGNITION FOR PORTABLE DEVICES WITH OPTIONAL POST PROCESSING

Patent №

US 7,324,943

Granted

2008-01-29

Filed 2003

Owner

MATSUSHITA ELECTRIC INDUSTRIAL CO., LTD.

Lab

AI components

3

ml · nlp · speech

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

10677174

A media capture device has an audio input receptive of user speech relating to a media capture activity in close temporal relation to the media capture activity. A plurality of focused speech recognition lexica respectively relating to media capture activities are stored on the device, and a speech recognizer recognizes the user speech based on a selected one of the focused speech recognition lexica. A media tagger tags captured media with generated speech recognition text, and a media annotator annotates the captured media with a sample of the user speech that is suitable for input to a speech recognizer. Tagging and annotating are based on close temporal relation between receipt of the user speech and capture of the captured media. Annotations may be converted to tags during post processing, employed to edit a lexicon using letter-to-sound rules and spelled word input, or matched directly to speech to retrieve captured media.

AI classification

Natural language1.00
Speech1.00
Machine learning0.93
AI hardware0.43
Knowledge representation0.10
Planning0.04
Vision0.01
Evolutionary computation0.00

Ownership

MATSUSHITA ELECTRIC INDUSTRIAL CO., LTD.

assignment · 145700315

Assignors

RIGAZIO, LUCA, BOMAN, ROBERT, NGUYEN, PATRICK, JUNQUA, JEAN-CLAUDE

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC