DocumentCode :
1070691
Title :
On the Fly Estimation of the Processes that Are Alive in an Asynchronous Message-Passing System
Author :
Mostefaoui, Achour ; Raynal, Michel ; Tredan, Gilles
Author_Institution :
IRISA, Univ. de Rennes 1, Rennes
Volume :
20
Issue :
6
fYear :
2009
fDate :
6/1/2009 12:00:00 AM
Firstpage :
778
Lastpage :
787
Abstract :
It is well known that in an asynchronous system where processes are prone to crash, it is impossible to design a protocol that provides each process with the set of processes that are currently alive. Basically, this comes from the fact that it is impossible to distinguish a crashed process from a process that is very slow or with which communications are very slow. Nevertheless, designing protocols that provide the processes with good approximations of the set of processes that are currently alive remains a real challenge in fault-tolerant-distributed computing. This paper proposes such a protocol, plus a second protocol that allows to cope with heterogeneous communication networks. These protocols consider a realistic computation model where the processes are provided with nonsynchronized local clocks and a function alpha () that takes a local duration Delta as a parameter, and returns an integer that is an estimate of the number of processes that could have crashed during that duration Delta. A simulation-based experimental evaluation of the proposed protocols is also presented. These experiments show that the protocols are practically relevant.
Keywords :
fault tolerant computing; message passing; protocols; system recovery; alive process estimation; approximation protocol design; asynchronous message-passing system; crash failure; fault-tolerant-distributed computing; heterogeneous communication network; process crash detection; Approximation protocol; asynchronous system; coverage assumption; crash detection; crash failure; fault-tolerance; message passing; nonsynchronized local clocks.;
fLanguage :
English
Journal_Title :
Parallel and Distributed Systems, IEEE Transactions on
Publisher :
ieee
ISSN :
1045-9219
Type :
jour
DOI :
10.1109/TPDS.2009.12
Filename :
4752815
Link To Document :
بازگشت