Visual Search Using Multiple Visual Input Modalities

Patent №

US 9,507,803

Granted

2016-11-29

Filed 2013

Owner

MICROSOFT CORPORATION

AI components

5

nlp · vision · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

14076890

Systems, methods, and computer-readable storage media for web-scale visual search capable of using a combination of visual input modalities are provided. An edgel index is created that includes shape-descriptors, including edgel-based representations, that correspond to each of a plurality of images. Each edgel-based representation includes pixels that depicts edges or boundary contours of an image and is created, at least in part, by segmenting the image into a plurality of image segments and performing a multi-phase contour detection on each segment. Upon receiving a search query having a visual query input, the visual query input is converted into shape-descriptors, including an edgel-based representation, and the shape-descriptors, including the edgel-based representation, of each of the plurality of images is compared with the shape-descriptors, including the edgel-based representation, of the visual query input to identify at least one image of the plurality of images that matches the visual query input.

Natural languageVisionKnowledge representationPlanningAI hardwareG06F 16/5854G06F 16/316G06F 16/3328G06F 16/532G06F 16/5838G06F 16/951

AI classification

Vision1.00
Knowledge representation0.97
AI hardware0.97
Natural language0.94
Planning0.54
Machine learning0.01
Evolutionary computation0.00
Speech0.00

Ownership

MICROSOFT CORPORATION

assignment · 321040159

© 2026 NYSGPT2525 LLC