Extraction of Data From Web Pages : A Vision Based Approach

Source: World Academy of Science, Engineering and Technology

Favorite

Free registration required

With the explosive growth of information sources available on the World Wide Web, it has become increasingly difficult to identify the relevant pieces of information, since web pages are often cluttered with irrelevant content like advertisements, navigation panels, copyright notices etc., surrounding the main content of the web page. Hence, tools for the mining of data regions, data records and data items need to be developed in order to provide value added services. Currently available automatic techniques to mine data regions from web pages are still unsatisfactory because of their poor performance and tag dependence.
Format:PDF Size:860.40
Date:Sep 2010