Browser

Extraction of Data From Web Pages : A Vision Based Approach

Free registration required

Executive Summary

With the explosive growth of information sources available on the World Wide Web, it has become increasingly difficult to identify the relevant pieces of information, since web pages are often cluttered with irrelevant content like advertisements, navigation panels, copyright notices etc., surrounding the main content of the web page. Hence, tools for the mining of data regions, data records and data items need to be developed in order to provide value added services. Currently available automatic techniques to mine data regions from web pages are still unsatisfactory because of their poor performance and tag dependence.

  • Format: PDF
  • Size: 860.4 KB