New Approaches in Engineering Research Vol. 1 | 2021

Semi-Automated Text Categorization Using Demonstration and Integration Based Term Set

 
 
 

Abstract


Manual Analysis of massive amounts of textual data requires incredible amount of processing time and effort in the interpretation of the text and organizing them in required format. In the current scenario, the major problem is with text or document categorization because of the high dimensionality of feature space. Now-a-days there are many methods available to deal with text feature selection. This paper aims at one such semi-automated text categorization feature selection methodology to deal with a enormous data using two phases of David Merrill’s First principles of instruction (FPI). It uses a pre-defined category group by providing them with the proper training set based on the demonstration and integration phase of FPI. The methodology involves the text tokenization, text categorization and text analysis.

Volume None
Pages None
DOI 10.9734/bpi/naer/v1/8766d
Language English
Journal New Approaches in Engineering Research Vol. 1

Full Text