2005•Unpublished venueRequires access

New Causal Message Logging Protocol with Asynchronous Checkpointing for Distributed Systems

Jinho Ahn

Open publisher page 1 citations

Abstract

Abstract. Causal message logging is an efficient approach for tolerating failures of processes in distributed systems because it has the advantages of both pessimistic and optimistic message logging approach. However, traditional causal message logging protocols prevent live processes from executing continuously their computation and require some synchronous logging to the stable storage during recovery. Although Elnozahy’s protocol solves the problems, it has the central recovery leader’s problem. Additionally, if it were integrated with asynchronous checkpointing, it may result in inconsistency problems in case of concurrent failures. In this paper, we present a new causal message logging protocol with asynchronous checkpointing to need to maintain only the latest checkpoint of each process and allow live processes to execute continuously their computation even in concurrent failures during recovery. Moreover, the protocol solves the problems of Elnozahy’s protocol and improves asynchrony during recovery because the protocol enables each recovering process to be responsible for only its recovery. 1

About this research paper

What this paper is about

Abstract. Causal message logging is an efficient approach for tolerating failures of processes in distributed systems because it has the advantages of both pessimistic and optimistic message logging approach. However, traditional causal message logging protocols prevent live processes from executing continuously their computation and require some synchronous logging to the stable storage during recovery. Although Elnozahy’s protocol solves the problems, it has the central recovery leader’s problem. Additionally, if it were integrated with asynchronous checkpointing, it may result in inconsistency problems in case of concurrent failures. In this paper, we present a new causal message logging protocol with asynchronous checkpointing to need to maintain only the latest checkpoint of each process and allow live processes to execute continuously their computation even in concurrent failures during recovery. Moreover, the protocol solves the problems of Elnozahy’s protocol and improves asynchrony during recovery because the protocol enables each recovering process to be responsible for only its recovery. 1

Why it matters

OpenAlex reports 1 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Abstract. Causal message logging is an efficient approach for tolerating failures of processes in distributed systems because it has the advantages of both pessimistic and optimistic message logging approach. However, traditional causal message logging protocols prevent live processes from executing continuously their computation and require some synchronous logging to the stable storage during recovery. Although Elnozahy’s protocol solves the problems, it has the central recovery leader’s problem. Additionally, if it were integrated with asynchronous checkpointing, it may result in inconsistency problems in case of concurrent failures. In this paper, we present a new causal message logging protocol with asynchronous checkpointing to need to maintain only the latest checkpoint of each process and allow live processes to execute continuously their computation even in concurrent failures during recovery. Moreover, the protocol solves the problems of Elnozahy’s protocol and improves asynchrony during recovery because the protocol enables each recovering process to be responsible for only its recovery. 1

Key concepts: Asynchronous communication, Computer science, Distributed computing, Protocol (science), Message passing, Two-phase commit protocol, Process (computing), Logging

Related papers

Back to paper searchBrowse research topicsOriginal source
New Causal Message Logging Protocol with Asynchronous Checkpointing for Distributed Systems — Research Paper | ScholarLens