Software

Extraction of Data From Web Pages : A Vision Based Approach

Date Added: Sep 2010
Format: PDF

With the explosive growth of information sources available on the World Wide Web, it has become increasingly difficult to identify the relevant pieces of information, since web pages are often cluttered with irrelevant content like advertisements, navigation panels, copyright notices etc., surrounding the main content of the web page. Hence, tools for the mining of data regions, data records and data items need to be developed in order to provide value added services. Currently available automatic techniques to mine data regions from web pages are still unsatisfactory because of their poor performance and tag dependence.