A NOVEL APPROACH ON TAMIL TEXT CLASSIFICATION USING C-FEATURE |
Author(s): |
| ARUNADEVI K , Coimbatore Institute of Technology, India; SAVEETH . R, Coimbatore Institute of Technology, India |
Keywords: |
| Tamil Text Classification, Feature Classification, Vocabulary set or bag-of–words, Text Mining, Natural Language Processing. |
Abstract |
|
Text Classification is one of the central issues in the information systems dealing with text data because of the increasing amount of information stored in a digital form. Text Classification Techniques have been applied on Tamil language to extract meaningful information and knowledge from unstructured Tamil text. Tamil language is a morphologically rich Dravidian language, classifying a Tamil document is different than classifying a English texts. In order to enhance the effectiveness of information extraction, we have compared feature extraction techniques and text classifiers and suggested an efficient C-feature (Compound feature), using C-feature we can create an efficient vocabulary set for Tamil text classification. |
Other Details |
|
Paper ID: IJSRDV2I5224 Published in: Volume : 2, Issue : 5 Publication Date: 01/08/2014 Page(s): 343-345 |
Article Preview |
|
|
|
|
