Patent №
US 6,563,952
Granted
2003-05-13
Filed 1999
Owner
HITACHI AMERICA LTD.
Lab
—
AI components
4
ml · nlp · vision · kr
Assignment
Recorded
Dataset
AIPD
2023_r1 edition
Application
09420252
The present invention is an apparatus and method for classifying high-dimensional sparse datasets. A raw data training set is flattened by converting it from categorical representation to a boolean representation. The flattened data is then used to build a class model on which new data not in the training set may be classified. In one embodiment, the class model takes the form of a decision tree, and large itemsets and cluster information are used as attributes for classification. In another embodiment, the class model is based on the nearest neighbors of the data to be classified. An advantage of the invention is that, by flattening the data, classification accuracy is increased by eliminating artificial ordering induced on the attributes. Another advantage is that the use of large itemsets and clustering increases classification accuracy.
AI classification
Ownership
HITACHI AMERICA LTD.
assignment · 107940233
Assignors
SRIVASTAVA, ANURAG, RAMKUMAR, G.D., SINGH, VINEET, RANKA, SANJAY
On an employer assignment, the assignors are typically the inventors.