2008•Unpublished venueRequires access

An Alternate Downloading Methodology of Webpages

Anirban Kundu, Alok Ranjan Pal, Tanay Sarkar, Moutan Banerjee, Subhendu Mandal, Rana Dattagupta, Debajyoti Mukhopadhyay

Open publisher page 1 citations

Abstract

We propose an advanced method for downloading Webpages from the Internet. In this technique, the whole system is considered as a bundle of crawlers which have been created dynamically at execution time. Numbers of crawlers are used depending on the requirement of downloading Webpages. The software module which interacts with WWW to search one or more Webpages is known as crawler. The numbers of crawlers are generated using the hierarchy structure of the Web server from which the data would be downloaded. Webpage downloader is an important issue for downloading Web documents from the Internet to facilitate a Web user in terms of knowledge gathering. This type of downloaders are very popular in the 'Information Technology' field. All kinds of public data, accessible throughout the world without any authentication, can be retrieved any time from any geographic location using the downloading methodology. Typically, a downloading technique has been utilized to accumulate Webpages of different domains within a single computer machine one at a time. So, our aim in this paper is to show an advanced technique for downloading a lot of related Webpages with a minimum effort and time using hierarchical downloader consisting of several dynamic crawlers.

About this research paper

What this paper is about

We propose an advanced method for downloading Webpages from the Internet. In this technique, the whole system is considered as a bundle of crawlers which have been created dynamically at execution time. Numbers of crawlers are used depending on the requirement of downloading Webpages. The software module which interacts with WWW to search one or more Webpages is known as crawler. The numbers of crawlers are generated using the hierarchy structure of the Web server from which the data would be downloaded. Webpage downloader is an important issue for downloading Web documents from the Internet to facilitate a Web user in terms of knowledge gathering. This type of downloaders are very popular in the 'Information Technology' field. All kinds of public data, accessible throughout the world without any authentication, can be retrieved any time from any geographic location using the downloading methodology. Typically, a downloading technique has been utilized to accumulate Webpages of different domains within a single computer machine one at a time. So, our aim in this paper is to show an advanced technique for downloading a lot of related Webpages with a minimum effort and time using hierarchical downloader consisting of several dynamic crawlers.

Why it matters

OpenAlex reports 1 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

We propose an advanced method for downloading Webpages from the Internet. In this technique, the whole system is considered as a bundle of crawlers which have been created dynamically at execution time. Numbers of crawlers are used depending on the requirement of downloading Webpages. The software module which interacts with WWW to search one or more Webpages is known as crawler. The numbers of crawlers are generated using the hierarchy structure of the Web server from which the data would be downloaded. Webpage downloader is an important issue for downloading Web documents from the Internet to facilitate a Web user in terms of knowledge gathering. This type of downloaders are very popular in the 'Information Technology' field. All kinds of public data, accessible throughout the world without any authentication, can be retrieved any time from any geographic location using the downloading methodology. Typically, a downloading technique has been utilized to accumulate Webpages of different domains within a single computer machine one at a time. So, our aim in this paper is to show an advanced technique for downloading a lot of related Webpages with a minimum effort and time using hierarchical downloader consisting of several dynamic crawlers.

Key concepts: Upload, Web page, Computer science, Web crawler, World Wide Web, The Internet, Web server, Static web page

Related papers

Back to paper searchBrowse research topicsOriginal source
An Alternate Downloading Methodology of Webpages — Research Paper | ScholarLens