NATURAL LANGUAGE DOMAIN CORPUS DATA SET CREATION BASED ON ENHANCED ROOT UTTERANCES

Patent №

US 11,664,010

Granted

2023-05-30

Filed 2020

Owner

FLORIDA POWER & LIGHT COMPANY

Lab

AI components

7

ml · nlp · vision · speech · kr · planning · hardware

Assignment

Recorded

Dataset

AIPD

2023_r1 edition

Application

17088071

Systems and methods for generating a natural language domain corpus to train a machine learning natural language understanding process. A base utterance expressing an intent and an intent profile indicating at least one of categories, keywords, concepts, sentiment, entities, or emotion of the intent are received. Machine translation translates the base utterance into a plurality of foreign language utterances and back into respective utterances in the target natural language to create a normalized utterance set. Analysis of each utterance in the normalized utterance set determines respective meta information for each such utterance. Comparison of the meta information to the intent profile determines a highest ranking matching utterance within the normalized utterance set. A set of natural language data to train a machine learning natural language understating process is created based on further natural language translations of the highest ranking matching utterance.

AI classification

Natural language1.00
Speech1.00
Machine learning1.00
Knowledge representation1.00
Planning1.00
Vision0.99
AI hardware0.53
Evolutionary computation0.00

Ownership

FLORIDA POWER & LIGHT COMPANY

assignment · 542570664

Assignors

MUSCHETT, BRIEN H., CALHOUN, JOSHUA D.

On an employer assignment, the assignors are typically the inventors.

© 2026 NYSGPT2525 LLC