An Advanced Approach on Focused Crawling with Anchor Text

Authors

  • S. Subatra Devi Hindustan College of Arts & Science, Chennai, Tamil Nadu, India.

DOI:

https://doi.org/10.9734/bpi/ctmcs/v3/2547F

Keywords:

Focused crawler, hyperlink, anchor text, Sibling, World Wide Web

Abstract

Title: An advanced approach with focused crawling for various anchor texts is discussed in this paper.

Background: Most of the search engines search the web with the anchor text to retrieve the relevant pages and answer the queries given by the users. The crawler usually searches the web pages and filters the unnecessary pages which can be done through focused crawling. A focused crawler generates its boundary to crawl the relevant pages based on the link and ignores the irrelevant pages on the web.

Methods and Findings: In this paper, an effective focused crawling method is implemented to improve the quality of the search. Here, three learning phases are considered namely, content-based, link-based and sibling-based learning are undergone to improve the navigation of the search. In this approach, the crawler crawls through the relevant pages efficiently and more relevant pages are retrieved in an effective way.

Study Objective: The objective of the study is that more number of relevant pages is retrieved for different anchor texts with three learning phases using focused crawling.

Published

2021-06-29

How to Cite

S. Subatra Devi. (2021). An Advanced Approach on Focused Crawling with Anchor Text. Current Topics on Mathematics and Computer Science Vol. 3, 22–35. https://doi.org/10.9734/bpi/ctmcs/v3/2547F