Language Modeling For Conversational Understanding Domains Using Semantic Web Resources

Patent №

US 9,679,558

Granted

2017-06-13

Filed 2014

Owner

MICROSOFT CORPORATION

AI components

6

ml · nlp · speech · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

14278659

Systems and methods are provided for training language models using in-domain-like data collected automatically from one or more data sources. The data sources (such as text data or user-interactional data) are mined for specific types of data, including data related to style, content, and probability of relevance, which are then used for language model training. In one embodiment, a language model is trained from features extracted from a knowledge graph modified into a probabilistic graph, where entity popularities are represented and the popularity information is obtained from data sources related to the knowledge. Embodiments of language models trained from this data are particularly suitable for domain-specific conversational understanding tasks where natural language is used, such as user interaction with a game console or a personal assistant application on personal device.

Machine learningNatural languageSpeechKnowledge representationPlanningAI hardwareG10L 15/063G06F 16/3329G06F 16/637G06F 40/40G10L 15/18G10L 15/183

AI classification

Natural language1.00
Machine learning1.00
Speech1.00
Knowledge representation0.98
AI hardware0.87
Planning0.59
Vision0.06
Evolutionary computation0.00

Ownership

MICROSOFT CORPORATION

assignment · 412890982

© 2026 NYSGPT2525 LLC