SLA-BASED JOB SUBMISSION AND SCHEDULING WITH THE GLOBUS TOOLKIT 4
Tenschert Alex, Kubert Roland
Abstract
Open-access reader
Tenschert Alex, Kubert Roland
Abstract
Open-access reader
High performance computing is nowadays mostly performed in a best effortfashion. This is surprising as the closely related topic of grid computing, whichdeals with the federation of resources from multiple domains in order to supportlarge jobs, and cloud computing, which promises seemingly infinite amounts ofcompute and storage, both offer quality of service (QoS), albeit in different ways.Long-term service level agreements (SLAs), which require the establishment ofSLAs long in advance of their actual usage, seem a promising way for the offeringof QoS guarantees in an HPC environment in a way that is not disruptive to thebusiness models employed today. This work uses the long-term SLA approachas a basis for the provisioning of service levels for HPC resources and presentsan SLA management framework to support this. Flexibility is provided byproviding SLAs with different service levels, support for which is integratedinto job submission and scheduling. The SLA management framework can, ona high level, be used in a generic fashion and an implementation is presentedthat is evaluated against a motivating scenario.
OpenAlex reports 2 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
High performance computing is nowadays mostly performed in a best effortfashion. This is surprising as the closely related topic of grid computing, whichdeals with the federation of resources from multiple domains in order to supportlarge jobs, and cloud computing, which promises seemingly infinite amounts ofcompute and storage, both offer quality of service (QoS), albeit in different ways.Long-term service level agreements (SLAs), which require the establishment ofSLAs long in advance of their actual usage, seem a promising way for the offeringof QoS guarantees in an HPC environment in a way that is not disruptive to thebusiness models employed today. This work uses the long-term SLA approachas a basis for the provisioning of service levels for HPC resources and presentsan SLA management framework to support this. Flexibility is provided byproviding SLAs with different service levels, support for which is integratedinto job submission and scheduling. The SLA management framework can, ona high level, be used in a generic fashion and an implementation is presentedthat is evaluated against a motivating scenario.
Key concepts: Computer science, Service-level agreement, Provisioning, Service level, Job scheduler, Cloud computing, Quality of service, Scheduling (production processes)