Design of Data Capture Program Based on Web Crawler Technology
Rui-xing CHEN, Yi Chen, Chunyuan Deng, Di Mo
Abstract
Rui-xing CHEN, Yi Chen, Chunyuan Deng, Di Mo
Abstract
With the advent of the Internet of things era and the vigorous development of electronic information era, the network information resources are growing exponentially. Faced with the demand of obtaining useful information, based on the general web crawler technology, this paper uses Python software to design a deep and optimized web crawler data fetching program. This crawler program can effectively solve a series of problems, such as waiting time, information overlapping and information incompleteness, so that the crawler has good performance.
OpenAlex reports 2 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
With the advent of the Internet of things era and the vigorous development of electronic information era, the network information resources are growing exponentially. Faced with the demand of obtaining useful information, based on the general web crawler technology, this paper uses Python software to design a deep and optimized web crawler data fetching program. This crawler program can effectively solve a series of problems, such as waiting time, information overlapping and information incompleteness, so that the crawler has good performance.
Key concepts: Web crawler, Computer science, Python (programming language), World Wide Web, Focused crawler, The Internet, Web application, Software