The failure detector abstraction

The failure detector abstraction
复制标题

DOI:
10.1145/1883612.1883616
复制
发表时间:
2011
期刊:
ACM Comput. Surv.
影响因子:
--
通讯作者:
F. Freiling;R. Guerraoui;P. Kuznetsov
F. Freiling;R. Guerraoui;P. Kuznetsov
中科院分区:
其他
文献类型:
--
作者:
F. Freiling;R. Guerraoui;P. Kuznetsov

文献摘要

被引文献

相似文献

故障检测器是分布式计算中的基本抽象。本文从两个维度考察了这一抽象概念。首先,我们将故障检测器作为构建块来研究,以简化可靠的分布式算法的设计。特别是,我们说明了故障检测器如何分解时序假设来检测分布式协议算法中的故障。其次,我们研究故障检测器作为可计算性基准。也就是说,我们调查最弱的故障检测器问题,并说明如何使用故障检测器对问题进行分类。我们还强调了故障检测器抽象在每个维度上的一些局限性。
A failure detector is a fundamental abstraction in distributed computing. This article surveys this abstraction through two dimensions. First we study failure detectors as building blocks to simplify the design of reliable distributed algorithms. In particular, we illustrate how failure detectors can factor out timing assumptions to detect failures in distributed agreement algorithms. Second, we study failure detectors as computability benchmarks. That is, we survey the weakest failure detector question and illustrate how failure detectors can be used to classify problems. We also highlight some limitations of the failure detector abstraction along each of the dimensions.