全部 标题 作者
关键词 摘要

OALib Journal期刊
ISSN: 2333-9721
费用:99美元

查看量下载量

相关文章

更多...
-  2017 

A Novel Approach on Focused Crawling with Anchor Text

DOI: 10.21767/2349-3917.100007

Keywords: Focused crawler, Hyperlink, Anchor text, Sibling, World wide web, list of open access journals, open access, open access journals, open access publication, open access publisher, open access publishing, open access journal articles, imedpub, imedpub publishing, insight medical publishing, imedpub online

Full-Text   Cite this paper   Add to My Lib

Abstract:

Title: A novel approach with focused crawling for various anchor texts is discussed in this paper. Background: Most of the search engines search the web with the anchor text to retrieve the relevant pages and answer the queries given by the users. The crawler usually searches the web pages and filters the unnecessary pages which can be done through focused crawling. A focused crawler generates its boundary to crawl the relevant pages based on the link and ignores the irrelevant pages on the web. Methods and findings: In this paper, an effective focused crawling method is implemented to improve the quality of the search. Here, three learning phases are considered namely, content-based, link-based and sibling-based learning are undergone to improve the navigation of the search. In this approach, the crawler crawls through the relevant pages efficiently and more relevant pages are retrieved in an effective way. Conclusion: It is proved experimentally that more number of relevant pages are retrieved for different anchor texts with three learning phases using focused crawling.

Full-Text

Contact Us

service@oalib.com

QQ:3279437679

WhatsApp +8615387084133