Web Scraping (Extraction of Deep Web Page Content) |
Author(s): |
| Santosh G N , The National Institute of Engineering,Mysore; S. Lokesh, The National Institute of Engineering,Mysore |
Keywords: |
| HTML, Web Scrapping |
Abstract |
|
Extracting useful information from the web is the most significant issue of concern for the realization of semantic web. This may be achieved by several ways among which Web Usage Mining, Web Scrapping and Semantic Annotation plays an important role. Web mining enables to find out the relevant results from the web and is used to extract meaningful information from the discovery patterns kept back in the servers. Web usage mining is a type of web mining which mines the information of access routes/manners of users visiting the websites. Web scraping, another technique, is a process of extracting useful information from HTML pages. The content is presented in a human-readable layout and is not intended to be processed by automatic systems. Therefore, it is necessary to separate the content in a web forum discussion from the layout before doing any further information mining. In this paper, explore and discuss some information extraction techniques on web like web usage mining, web scrapping for a better or efficient information extraction on the web illustrated with examples |
Other Details |
|
Paper ID: IJSRDV3I21034 Published in: Volume : 3, Issue : 2 Publication Date: 01/05/2015 Page(s): 2403-2405 |
Article Preview |
|
|
|
|
