Article,

Semi Automated Text Categorization Using Demonstration Based Term Set

.
International Journal of Computer Science, Engineering and Applications (IJCSEA), 02 (04): 71-77 (August 2012)
DOI: 10.5121/ijcsea.2012.2408

Abstract

Manual Analysis of huge amount of textual data requires a tremendous amount of processing time and effort in reading the text and organizing them in required format. In the current scenario, the major problem is with text categorization because of the high dimensionality of feature space. Now-a-days there are many methods available to deal with text feature selection. This paper aims at such semi automated text categorization feature selection methodology to deal with a massive data using one of the phases of David Merrill’s First principles of instruction (FPI). It uses a pre-defined category group by providing them with the proper training set based on the demonstration phase of FPI. The methodology involves the text tokenization, text categorization and text analysis.

Tags

Users

  • @ijcsea

Comments and Reviews