Gang scheduling in a distributed system under processor failures and time-varying gang size
Helen D. Karatza
Abstract
Helen D. Karatza
Abstract
In this paper we study the performance of a distributed system which is subject to hardware failures and subsequent repairs. A special type of scheduling called gang scheduling is considered, under which jobs consist of a number of interacting tasks which are scheduled to run simultaneously on distinct processors. System performance is examined and compared in cases where different distributions for the number of parallel tasks per job (gang size) are employed We examine cases where gang size is defined by a specific distribution and also a case where gang size distribution varies with time.
OpenAlex reports 11 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
In this paper we study the performance of a distributed system which is subject to hardware failures and subsequent repairs. A special type of scheduling called gang scheduling is considered, under which jobs consist of a number of interacting tasks which are scheduled to run simultaneously on distinct processors. System performance is examined and compared in cases where different distributions for the number of parallel tasks per job (gang size) are employed We examine cases where gang size is defined by a specific distribution and also a case where gang size distribution varies with time.
Key concepts: Scheduling (production processes), Computer science, Processor scheduling, Execution time, Parallel computing, Distributed computing, Operating system, Mathematical optimization