article

Web Scraping (Extraction of Deep Web Page Content)

  • IJSRD : international journal for scientific research and development
  • IJRSD
Research footprint

At a glance

Citations
0
References
6
Comments
0
Paper overview

Abstract

Extracting useful information from the web is the most significant issue of concern for the realization of semantic web. This may be achieved by several ways among which Web Usage Mining, Web Scrapping and Semantic Annotation plays an important role. Web mining enables to find out the relevant results from the web and is used to extract meaningful information from the discovery patterns kept back in the servers. Web usage mining is a type of web mining which mines the information of access routes/manners of users visiting the websites. Web scraping, another technique, is a process of extracting useful information from HTML pages. The content is presented in a human-readable layout and is not intended to be processed by automatic systems. Therefore, it is necessary to separate the content in a web forum discussion from the layout before doing any further information mining. In this paper, explore and discuss some information extraction techniques on web like web usage mining, web scrapping for a better or efficient information extraction on the web illustrated with examples

Record transparency

Publication details

OpenAlex
W2248775472
Document type
article
Language
EN
Source
IJSRD : international journal for scientific research and development
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.