Text Classification using the Concept of Association Rule of Data Mining

As the amount of online text increases, the demand for text classification to aid the analysis and management of text is increasing. Text is cheap, but information, in the form of knowing what classes a text belongs to, is expensive. Automatic classification of text can provide this information at low cost, but the classifiers themselves must be built with expensive human effort, or trained from texts which have themselves been manually classified. In this paper we will discuss a procedure of classifying text using the concept of association rule of data mining. Association rule mining technique has been used to derive feature set from pre-classified text documents. Naive Bayes classifier is then used on derived features for final classification.

Paper

References (10)

04“Data Mining: Practical Machine Learning Larning Tools and Techniques with Java Implementation”2000
05Feature Selection and Feature Extraction for Text Categorization, appeared in Speech and Natural Language1992 · Proceedings of a workshop held at Harriman
06To discover the set of frequent 2-itemsets
07The transactions D are scanned and the support count of each candidate itemset in C2 is accumulated
08P({based,system}½CS)=0.02, P({based,system}½EE)=0.0067, P({based,system}½ME)
09removing unnecessary wordsour thesis work
10Automatic Keyphrase ExtractionAutomatic Keyphrase Extraction

Similar papers

© 2026 NYSGPT2525 LLC