TECHNIQUES FOR DISAMBIGUATING SPEECH INPUT USING MULTIMODAL INTERFACES

Patent №

US 7,684,985

Granted

2010-03-23

Filed 2003

Owner

KIRUSA, INC.

Lab

AI components

3

nlp · speech · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

10733793

A technique is disclosed for disambiguating speech input for multimodal systems by using a combination of speech and visual I/O interfaces. When the user's speech input is not recognized with sufficiently high confidence, a the user is presented with a set of possible matches using a visual display and/or speech output. The user then selects the intended input from the list of matches via one or more available input mechanisms (e.g., stylus, buttons, keyboard, mouse, or speech input). These techniques involve the combined use of speech and visual interfaces to correctly identify user's speech input. The techniques disclosed herein may be utilized in computer devices such as PDAs, cellphones, desktop and laptop computers, tablet PCs, etc.

AI classification

Speech1.00
Natural language1.00
AI hardware0.88
Machine learning0.03
Knowledge representation0.01
Vision0.01
Planning0.00
Evolutionary computation0.00

Ownership

KIRUSA, INC.

assignment · 153080895

Assignors

DOMINACH, RICHARD, ISUKAPALLI, SASTRY, SIBAL, SANDEEP, VAIDYA, SHIRISH

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC