INFORMATION RETRIEVAL UTILIZING SEMANTIC REPRESENTATION OF TEXT

Patent №

US 6,076,051

Granted

2000-06-13

Filed 1997

Owner

MICROSOFT CORPORATION

AI components

4

ml · nlp · kr · planning

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

08886814

The present invention is directed to performing information retrieval utilizing semantic representation of text. In a preferred embodiment, a tokenizer generates from an input string information retrieval tokens that characterize the semantic relationship expressed in the input string. The tokenizer first creates from the input string a primary logical form characterizing a semantic relationship between selected words in the input string. The tokenizer then identifies hypernyms that each have an "is a" relationship with one of the selected words in the input string. The tokenizer then constructs from the primary logical form one or more alternative logical forms. The tokenizer constructs each alternative logical form by, for each of one or more of the selected words in the input string, replacing the selected word in the primary logical form with an identified hypernym of the selected word. Finally, the tokenizer generates tokens representing both the primary logical form and the alternative logical forms. The tokenizer is preferably used to generate tokens for both constructing an index representing target documents and processing a query against that index.

Machine learningNatural languageKnowledge representationPlanningG06F 40/30G06F 16/3344G06F 40/211G06F 40/284Y10S 707/99932Y10S 707/99935

AI classification

Natural language1.00
Knowledge representation1.00
Machine learning0.99
Planning0.97
Vision0.37
AI hardware0.06
Speech0.01
Evolutionary computation0.00

Ownership

MICROSOFT CORPORATION

assignment · 86670309

Assignors

MESSERLY,, JOHN J., HEIDORN, GEORGE E., RICHARDSON, STEPHEN D., DOLAN, WILLIAM B., JENSEN, KAREN

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC