Language Modeling For Conversational Understanding Domains Using Semantic Web Resources
Patent №
US 9,679,558
Granted
2017-06-13
Filed 2014
Owner
MICROSOFT CORPORATION
Lab
AI components
6
ml · nlp · speech · kr · planning · hardware
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
14278659
Systems and methods are provided for training language models using in-domain-like data collected automatically from one or more data sources. The data sources (such as text data or user-interactional data) are mined for specific types of data, including data related to style, content, and probability of relevance, which are then used for language model training. In one embodiment, a language model is trained from features extracted from a knowledge graph modified into a probabilistic graph, where entity popularities are represented and the popularity information is obtained from data sources related to the knowledge. Embodiments of language models trained from this data are particularly suitable for domain-specific conversational understanding tasks where natural language is used, such as user interaction with a game console or a personal assistant application on personal device.
AI classification
Ownership
MICROSOFT CORPORATION
assignment · 412890982