The Design and Implement of High Efficient Incremental Microblogging Crawler
Dayong Shen, Hui Wang, Cao Jianping, Pei Li, Zhi‐Hong Jiang
Abstract
Dayong Shen, Hui Wang, Cao Jianping, Pei Li, Zhi‐Hong Jiang
Abstract
With the rapid development of microblog technology, many interesting research issues on microblog have aroused growing attention. Data fetching from microblog is the groundwork of these researches. In this paper we take Sina microblog (also called Weibo) as the crawling site, designing and implementing a high efficient incremental microblog crawler based on the classic multi-producers and multi-consumers model. Experimental results demonstrate that the crawler can collect real time microblog information efficiently and precisely.
A significance statement is not available in the OpenAlex record.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
With the rapid development of microblog technology, many interesting research issues on microblog have aroused growing attention. Data fetching from microblog is the groundwork of these researches. In this paper we take Sina microblog (also called Weibo) as the crawling site, designing and implementing a high efficient incremental microblog crawler based on the classic multi-producers and multi-consumers model. Experimental results demonstrate that the crawler can collect real time microblog information efficiently and precisely.
Key concepts: Microblogging, Web crawler, Crawling, Social media, Computer science, Focused crawler, World Wide Web, Data science