High Impact Factor : 4.396 icon | Submit Manuscript Online icon |

World Wide Web Information Retrieval Using Clustering

Author(s):

Manjunath Naik , RNSIT; Sudha V, RNSIT

Keywords:

Crowdsourcing, Clustering, Preprocessing, SVM .

Abstract

Crowdsourcing, a distributed process that involves outsourcing tasks to a network of people, is increasingly used by companies for generating solutions to problems of various kinds. In this way, thousands of people contribute a large amount of text data that needs to already be structured during the process of idea generation in order to avoid repetitions and to maximize the solution space. This is a hard information retrieval problem as the texts are very short and have little predefined structure. We present a solution that involves three steps: text data preprocessing, clustering, and visualization. In this contribution, we focus on clustering and visualization by presenting a Support Vector machine approach that is able to learn the principal components of the data while the data set is continuously growing in size. We compare our approach to standard clustering applications and demonstrate its superiority with respect to classification reliability on a real-world example.

Other Details

Paper ID: IJSRDV2I4181
Published in: Volume : 2, Issue : 4
Publication Date: 01/07/2014
Page(s): 351-355

Article Preview

Download Article