An Automatic and Scalable Application Crawler for Large-Scale Mobile Internet Content Retrieval
Mingyi Huang, Yongqiang Lyu, Hao Yin
Abstract
Open-access reader
Mingyi Huang, Yongqiang Lyu, Hao Yin
Abstract
Open-access reader
The mobile internet has grown ubiquitous across the globe with the widespread use of smart devices.However, the designs of modern mobile operating systems and their applications limit content retrieval with mobile applications.The mobile internet is not as accessible as the traditional web, having more man-made restrictions and lacking a unified approach for crawling and content retrieval.In this study, we propose an automatic and scalable mobile application content crawler, which can recognize the interaction paths of mobile applications, representing them as interaction graphs and automatically collecting content according to the graphs in a parallel manner.The crawler was verified by retrieving content from 50 non-game applications from the Google Play Store using the Android platform.The experiment showed the efficiency and scalability potential of our crawler for large-scale mobile internet content retrieval.
OpenAlex reports 1 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
The mobile internet has grown ubiquitous across the globe with the widespread use of smart devices.However, the designs of modern mobile operating systems and their applications limit content retrieval with mobile applications.The mobile internet is not as accessible as the traditional web, having more man-made restrictions and lacking a unified approach for crawling and content retrieval.In this study, we propose an automatic and scalable mobile application content crawler, which can recognize the interaction paths of mobile applications, representing them as interaction graphs and automatically collecting content according to the graphs in a parallel manner.The crawler was verified by retrieving content from 50 non-game applications from the Google Play Store using the Android platform.The experiment showed the efficiency and scalability potential of our crawler for large-scale mobile internet content retrieval.
Key concepts: Computer science, Web crawler, Scalability, Information retrieval, The Internet, Scale (ratio), World Wide Web, Mobile internet