VOICE TAGGING, VOICE ANNOTATION, AND SPEECH RECOGNITION FOR PORTABLE DEVICES WITH OPTIONAL POST PROCESSING
Patent №
US 7,324,943
Granted
2008-01-29
Filed 2003
Owner
MATSUSHITA ELECTRIC INDUSTRIAL CO., LTD.
Lab
—
AI components
3
ml · nlp · speech
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
10677174
A media capture device has an audio input receptive of user speech relating to a media capture activity in close temporal relation to the media capture activity. A plurality of focused speech recognition lexica respectively relating to media capture activities are stored on the device, and a speech recognizer recognizes the user speech based on a selected one of the focused speech recognition lexica. A media tagger tags captured media with generated speech recognition text, and a media annotator annotates the captured media with a sample of the user speech that is suitable for input to a speech recognizer. Tagging and annotating are based on close temporal relation between receipt of the user speech and capture of the captured media. Annotations may be converted to tags during post processing, employed to edit a lexicon using letter-to-sound rules and spelled word input, or matched directly to speech to retrieve captured media.
AI classification
Ownership
MATSUSHITA ELECTRIC INDUSTRIAL CO., LTD.
assignment · 145700315
Assignors
RIGAZIO, LUCA, BOMAN, ROBERT, NGUYEN, PATRICK, JUNQUA, JEAN-CLAUDE
On an employer assignment, the assignors are typically the inventors.