2015Unpublished venueRequires access

An Overview of Approaches Used In Focused Crawlers

Parigha V. Suryawanshi

Open publisher page 0 citations

Abstract

Web is a repository n where there is variety of information available provided by millions of web content providers. Numerous WebPages are added to web every day and the content keeps changing. Search engines are used to mine this information and the most important part of search engine is a web crawler also known as web spider. A web crawler basically is software that crawls or browses the WebPages in the World Wide Web. There are many types of crawlers having different methods of crawling like parallel crawler, distributed crawler, focused crawler, parallel crawler, and incremental crawler. In recent years, focused crawling has attracted considerable interest in research due to the increasing need of digital libraries and domain-specific search engines. This paper reviews different researches done in focused crawler which are also known as topic specific web crawler.

About this research paper

What this paper is about

Web is a repository n where there is variety of information available provided by millions of web content providers. Numerous WebPages are added to web every day and the content keeps changing. Search engines are used to mine this information and the most important part of search engine is a web crawler also known as web spider. A web crawler basically is software that crawls or browses the WebPages in the World Wide Web. There are many types of crawlers having different methods of crawling like parallel crawler, distributed crawler, focused crawler, parallel crawler, and incremental crawler. In recent years, focused crawling has attracted considerable interest in research due to the increasing need of digital libraries and domain-specific search engines. This paper reviews different researches done in focused crawler which are also known as topic specific web crawler.

Why it matters

A significance statement is not available in the OpenAlex record.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Web is a repository n where there is variety of information available provided by millions of web content providers. Numerous WebPages are added to web every day and the content keeps changing. Search engines are used to mine this information and the most important part of search engine is a web crawler also known as web spider. A web crawler basically is software that crawls or browses the WebPages in the World Wide Web. There are many types of crawlers having different methods of crawling like parallel crawler, distributed crawler, focused crawler, parallel crawler, and incremental crawler. In recent years, focused crawling has attracted considerable interest in research due to the increasing need of digital libraries and domain-specific search engines. This paper reviews different researches done in focused crawler which are also known as topic specific web crawler.

Key concepts: Web crawler, Focused crawler, World Wide Web, Computer science, Web search engine, Crawling, Web page, Information retrieval

Related papers

Back to paper searchBrowse research topicsOriginal source
An Overview of Approaches Used In Focused Crawlers — Research Paper | ScholarLens