An Alternate Downloading Methodology of Webpages
Anirban Kundu, Alok Ranjan Pal, Tanay Sarkar, Moutan Banerjee, Subhendu Mandal, Rana Dattagupta, Debajyoti Mukhopadhyay
Abstract
Anirban Kundu, Alok Ranjan Pal, Tanay Sarkar, Moutan Banerjee, Subhendu Mandal, Rana Dattagupta, Debajyoti Mukhopadhyay
Abstract
We propose an advanced method for downloading Webpages from the Internet. In this technique, the whole system is considered as a bundle of crawlers which have been created dynamically at execution time. Numbers of crawlers are used depending on the requirement of downloading Webpages. The software module which interacts with WWW to search one or more Webpages is known as crawler. The numbers of crawlers are generated using the hierarchy structure of the Web server from which the data would be downloaded. Webpage downloader is an important issue for downloading Web documents from the Internet to facilitate a Web user in terms of knowledge gathering. This type of downloaders are very popular in the 'Information Technology' field. All kinds of public data, accessible throughout the world without any authentication, can be retrieved any time from any geographic location using the downloading methodology. Typically, a downloading technique has been utilized to accumulate Webpages of different domains within a single computer machine one at a time. So, our aim in this paper is to show an advanced technique for downloading a lot of related Webpages with a minimum effort and time using hierarchical downloader consisting of several dynamic crawlers.
OpenAlex reports 1 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
We propose an advanced method for downloading Webpages from the Internet. In this technique, the whole system is considered as a bundle of crawlers which have been created dynamically at execution time. Numbers of crawlers are used depending on the requirement of downloading Webpages. The software module which interacts with WWW to search one or more Webpages is known as crawler. The numbers of crawlers are generated using the hierarchy structure of the Web server from which the data would be downloaded. Webpage downloader is an important issue for downloading Web documents from the Internet to facilitate a Web user in terms of knowledge gathering. This type of downloaders are very popular in the 'Information Technology' field. All kinds of public data, accessible throughout the world without any authentication, can be retrieved any time from any geographic location using the downloading methodology. Typically, a downloading technique has been utilized to accumulate Webpages of different domains within a single computer machine one at a time. So, our aim in this paper is to show an advanced technique for downloading a lot of related Webpages with a minimum effort and time using hierarchical downloader consisting of several dynamic crawlers.
Key concepts: Upload, Web page, Computer science, Web crawler, World Wide Web, The Internet, Web server, Static web page